Posted On: Aug 19, 2026
Job Type
Experience
5 - 10 Years
Salary
Depends on Experience
Work Arrangement
Travel Requirement
0%
Required Skills
We are seeking a Senior Observability Engineer with strong experience designing, implementing, and managing enterprise-scale observability platforms. The ideal candidate will have hands-on expertise with Grafana, Terraform, OpenTelemetry, monitoring, logging, alerting, and distributed tracing, along with experience migrating from legacy observability and monitoring platforms.
Design, implement, and maintain enterprise-scale observability and monitoring solutions.
Administer and manage Grafana across large-scale enterprise environments.
Develop and maintain infrastructure using Terraform and Infrastructure as Code (IaC) practices.
Implement and optimize monitoring, alerting, logging, and distributed tracing solutions.
Design and manage OpenTelemetry instrumentation and telemetry pipelines.
Support migrations from platforms such as Splunk, Dynatrace, AppDynamics, New Relic, OpenText OBM, or similar monitoring and observability tools.
Develop dashboards, alerts, metrics, logs, and traces to improve system visibility and operational reliability.
Troubleshoot observability infrastructure and resolve monitoring and telemetry-related issues.
Collaborate with Platform Engineering, DevOps, SRE, Application Development, and Operations teams to establish effective observability standards.
Automate observability platform deployment, configuration, and management using IaC and automation tools.
Establish best practices for scalability, reliability, security, and performance of observability platforms.
Document observability architectures, configurations, standards, and operational procedures.
5+ years of experience in observability, monitoring, operations, SRE, DevOps, or platform engineering.
Hands-on experience administering Grafana in large-scale enterprise environments.
Strong experience with Terraform and Infrastructure as Code (IaC).
Hands-on experience with OpenTelemetry, including instrumentation concepts and telemetry pipelines.
Experience implementing and managing monitoring, alerting, logging, and distributed tracing solutions.
Experience migrating from observability and monitoring platforms such as Splunk, Dynatrace, AppDynamics, New Relic, OpenText OBM, or similar tools.
Strong understanding of enterprise observability architecture and modern telemetry practices.
Experience troubleshooting complex monitoring and observability issues in production environments.
Strong analytical, problem-solving, communication, and collaboration skills.
Experience with cloud platforms such as AWS, Azure, or GCP.
Experience with Kubernetes and containerized environments.
Familiarity with Prometheus, Loki, Tempo, Elasticsearch, or similar observability technologies.
Experience with CI/CD pipelines and automation.
Experience establishing observability standards and best practices across large enterprise environments.
Job ID: 2C322094
Posted By
Shayne sha
Sr. Recruiter