Senior Dynatrace Observability Engineer
meta black- Posted 9 hours ago
- Be among the first 10 applicants
Job Description
Experience: 8-12+ Years | Location: Onsite - any Cognizant office / Hybrid
Key Responsibilities
• Design, implement, and manage Dynatrace-based enterprise observability solutions across applications, infrastructure, databases, APIs, containers, Kubernetes, Docker, and cloud environments.
• Configure and administer OneAgent, ActiveGate, dashboards, management zones, alerts, monitoring configurations, custom metrics, SLOs, SLIs, and service-level monitoring.
• Perform APM, distributed tracing, log, metric, event, PurePath, and Service Flow analysis; use Davis AI to identify anomalies, dependencies, and root causes.
• Lead RCA for complex performance issues and support production releases, incident management, capacity planning, performance optimization, reliability, availability, and MTTR improvement.
• Integrate Dynatrace with ServiceNow, Splunk, PagerDuty, CI/CD tools, and enterprise platforms; develop automation through Dynatrace APIs, workflows, Python, PowerShell, or shell scripts.
• Collaborate with DevOps, SRE, development, infrastructure, and application support teams; establish observability standards, best practices, governance, and mentor junior engineers.
Required Skills
• 8-12+ years in IT Operations, APM, Observability, SRE, or Performance Engineering with enterprisescale Dynatrace implementation experience.
• Deep knowledge of Dynatrace OneAgent, ActiveGate, Davis AI, distributed tracing, Service Flow, PurePath, dashboards, alerting, logs, metrics, traces, and events.
• Strong experience with Kubernetes, Docker, containers, microservices, API-based architectures, and AWS, Azure, or GCP environments.
Strong troubleshooting, performance analysis, RCA, REST API, automation/scripting, DevOps, CI/CD, Agile, and SRE capabilities.
Preferred / Nice-to-Have Skills
• Dynatrace Certified Professional/Associate; experience with OpenTelemetry and cloud-native observability.
• Knowledge of Terraform, Ansible, Jenkins, GitLab, or Azure DevOps; experience with ITSM/SIEM integrations such as ServiceNow, Splunk, or PagerDuty.
• Experience defining and implementing SLIs, SLOs, SLAs, error budgets, enterprise observability strategy, governance, and platform adoption.
Key Skills
Dynatrace | Observability | APM | OneAgent | ActiveGate | Davis AI | PurePath | Distributed Tracing | Service Flow | Kubernetes | Docker | Microservices | AWS | Azure | GCP | OpenTelemetry | SRE | DevOps | CI/CD | SLO | SLI | RCA | Performance Engineering
More Info
Key Skills
Davis AI
OneAgent
PurePath
CI CD
Service Flow
SLI
ActiveGate
OpenTelemetry
SLO
Distributed Tracing
Observability
