Chennai, Tamil Nadu
Job Summary
We are seeking a highly skilled Senior Observability Engineer to lead the evolution of our enterprise observability platform. In this role, you will transition our systems to an open, standards-based observability architecture using OpenTelemetry (OTel). You will design and scale OTel solutions across polyglot microservices (Java, Python, etc.) and Google Cloud Platform (GCP) resources, routing telemetry to Dynatrace, Prometheus, and Jaeger. A major focus of this role is improving the overall developer experience by creating onboarding tools, documentation, migration paths, and automated Quality of Service (QoS) reporting.
Key Responsibilities
Expand Observability Coverage: Design and implement standardized OpenTelemetry (OTel) SDK and API configurations for polyglot environments, focusing on Python and Java. Architect and deploy OTel-based collection mechanisms for GCP-native resources, including serverless components (Cloud Functions) and asynchronous messaging (Pub/Sub). Ensure all OTel-ingested metrics, traces, and logs are seamlessly mapped, enriched, and visualized within Dynatrace using native OTLP ingestion. Enhance Interoperability: Design, deploy, and maintain high-availability OpenTelemetry Collector pipelines to receive, process, batch, and export telemetry data. Configure OTel pipelines to dynamically route telemetry to multiple backends (e.g., exporting traces to Dynatrace and Jaeger, and metrics to Prometheus). Establish unified semantic conventions, tagging schemas, and context propagation standards across all services. Improve User Experience & Developer Enablement: Build a frictionless onboarding experience for software engineering teams by creating reusable OTel templates and bootstrap libraries. Create automated migration tools and scripts to help teams transition from legacy/proprietary agents to OpenTelemetry. Design and build automated pipelines to generate customizable Quality of Service (QoS) and Service Level Objective (SLO) reports using Dynatrace and OTel APIs.
Skill Requirements
Required Technical Skills & Qualifications Primary Technical Skills: Cloud Platform: Strong hands-on experience with Google Cloud Platform (GCP), specifically Google Compute Engine (GCE) and Google Cloud Storage (GCS). Infrastructure as Code: Proven expertise in Terraform for automated provisioning of GCP resources. Backend Development: Strong proficiency in Python and hands-on experience with the Flask framework. Frontend Development: Solid understanding of UI/UX principles with hands-on development experience in React JS. API Management: Practical knowledge of Apigee for API gateway configuration, security, and management. Developer Experience & Tools: Developer Portal: Familiarity or hands-on experience with Red Hat Developer Hub (RHDH) or Backstage. Version Control: Strong command of Git (GitHub, GitLab, or Bitbucket). Collaboration & ITSM: Practical experience using Jira for agile project management and ServiceNow for service requests, incident management, and change control.
Other Requirements
Google Cloud Certified Associate Cloud Engineer or Professional Cloud Architect. HashiCorp Certified: Terraform Associate. Experience building developer self-service portals or automated provisioning pipelines. Understanding of containerization (Docker, Google Kubernetes Engine) is a plus.
We are seeking a highly skilled Senior Observability Engineer to lead the evolution of our enterprise observability platform. In this role, you will transition our systems to an open, standards-based observability architecture using OpenTelemetry (OTel). You will design and scale OTel solutions across polyglot microservices (Java, Python, etc.) and Google Cloud Platform (GCP) resources, routing telemetry to Dynatrace, Prometheus, and Jaeger
#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply-#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply-