Required Skills: Grafana Dashboard Development & Administration, Grafana Alerting & Notification Management, Prometheus, Loki, Tempo, OpenTelemetry, Elasticsearch, OpenSearch, AWS CloudWatch, Azure, GCP, Kubernetes, Container Monitoring, CI/CD, DevOps, Terraform, Ansible, Python, Bash, Shell, PowerShell, Metrics, Logs & Distributed Tracing
Job Description
We are looking for an experienced Grafana Engineer(Full-time) to design, implement, and manage enterprise-grade monitoring and observability solutions. If you have expertise in Grafana, Prometheus, Loki, OpenTelemetry, Kubernetes, and AWS Cloud, we'd love to hear from you!
Key Responsibilities
- Design and maintain Grafana dashboards, visualizations, and reports.
- Configure Grafana Alerting for proactive monitoring and incident management.
- Integrate Grafana with Prometheus, Loki, Elasticsearch/OpenSearch, SQL databases, AWS CloudWatch, Azure Monitor, and GCP Operations.
- Build observability solutions covering metrics, logs, and distributed tracing.
- Monitor Kubernetes, containers, cloud infrastructure, and enterprise applications.
- Collaborate with DevOps, SRE, Infrastructure, and Platform Engineering teams.
- Automate dashboard deployment using Terraform, Ansible, or other IaC tools.
- Perform root cause analysis and optimize monitoring systems.
- Create technical documentation, operational runbooks, and monitoring standards.
Required Skills
✅ Strong experience with Grafana Dashboard Development & Administration
✅ Grafana Alerting & Notification Management
✅ Prometheus, Loki, Tempo, OpenTelemetry
✅ Elasticsearch/OpenSearch
✅ AWS CloudWatch (Azure & GCP knowledge is a plus)
✅ Kubernetes & Container Monitoring
✅ CI/CD & DevOps Practices
✅ Terraform / Ansible
✅ Python, Bash/Shell, or PowerShell
✅ Metrics, Logs & Distributed Tracing
Preferred Skills
- Grafana Enterprise
- OpenTelemetry Implementation
- AIOps & Observability Platforms
- Splunk
- Dynatrace
- Datadog
- AppDynamics
- New Relic