s
shahawal1

Shah Khair

@shahawal1

Observability Engineer Grafana, Alloy, Prometheus, Loki, Tempo, Otel, Kubernetes

Pakistan
Engels, Urdu, Hindi
Sommige informatie wordt in het Engels weergegeven.
Over mij
I’m an Observability Engineer specializing in Grafana, Prometheus, Loki, Tempo, OpenTelemetry and Kubernetes. I design and deploy production-grade, open-source observability platforms that replace costly SaaS tools like Datadog, delivering comparable—and often more flexible—logs, metrics, traces, dashboards and alerting. I’ve migrated production environments to Grafana-based observability, optimized retention and storage, and delivered significant cost savings while maintaining comprehensive visibility across microservices and Kubernetes.... Lees meer

Skills

s
shahawal1
Shah Khair
offline • 
Gemiddelde reactietijd: 2 uur

Bekijk mijn diensten

Programmering en technologie
I will migrate your datadog observability to grafana lgtm on k8s

Werkervaring

LogicEra

Observability and Site Reliability Engineer

LogicEra • Freelance

May 2024 - Present2 yrs 3 mos

Observability Engineer and DevSecOps practitioner with hands-on production experience designing and operating full-stack observability platforms across Azure and AWS. Experienced in Kubernetes monitoring, metrics, centralized logging, distributed tracing, dashboards, alerting, and OpenTelemetry. Led the migration from Datadog SaaS to a self-hosted Grafana-based observability stack, delivering comparable observability capabilities while reducing annual costs by approximately $72,000. Skilled in Prometheus, Grafana, Loki, Tempo, Alloy, Alertmanager, ELK Stack, Kubernetes, CI/CD, and cloud platforms. Strong focus on reliability, cost optimization, automation, and practical production troubleshooting. KEY ACHIEVEMENT Observability Platform Migration and Cost Optimization - Approximately $72,000 Annual Savings Identified high recurring Datadog SaaS licensing costs and designed and deployed a production-grade, self-hosted Grafana observability platform on Kubernetes. Implemented metrics, centralized logging, distributed tracing, dashboards, alerting, and telemetry collection using Prometheus, Grafana, Loki, Tempo, Alloy, OpenTelemetry, and Alertmanager. Replaced the previous SaaS-based monitoring approach, improved data ownership and visibility, eliminated unpredictable licensing costs, and maintained reliable observability across production microservices.