Search Jobs

Search by job, company or skills

Observability & Sr. Observability Platform Engineer

Observability & Sr. Observability Platform Engineer

American Express Global Business Travel
  • Posted 5 days ago
  • Be among the first 10 applicants

Job Description

What You'll Do

  • Architect and lead the design of comprehensive observability platforms demonstrating ELK Stack, New Relic, Datadog, Alertsite, and emerging technologies
  • Establish observability standards, guidelines, and governance frameworks across the organization
  • Mentor and guide junior engineers and platform teams in observability implementation and optimization
  • Develop advanced monitoring strategies, custom dashboards, and intelligent alerting frameworks
  • Lead multi-functional initiatives to instrument applications and infrastructure for end-to-end visibility
  • Optimize observability infrastructure for scalability, cost-efficiency, and performance at enterprise scale
  • Present technical insights and recommendations to senior leadership and stakeholders

What We're Looking For

  • 8+ years of experience in platform engineering, DevOps, systems engineering, or related roles
  • 6+ years of hands-on expertise with enterprise observability platforms (ELK Stack, New Relic, Datadog, Alertsite, or similar)
  • Deep understanding of monitoring architectures, logging strategies, metrics collection, distributed tracing, and alerting frameworks
  • Strong hands-on experience implementing synthetic monitoring and end-to-end transaction monitoring
  • Deep expertise in Application Performance Monitoring (APM) concepts, implementation, and optimization
  • Extensive experience with Real User Monitoring (RUM) and digital/browser/mobile app observability
  • Expertise in defining, measuring, and tracking MTTA, MTTR, MTTD, and other incident metrics
  • Advanced proficiency in scripting and programming languages (Python, Go, Bash, or similar)
  • Extensive experience with cloud platforms (AWS, Azure, or GCP) and multi-cloud environments

Preferred Qualifications

  • Experience architecting observability solutions for large-scale, distributed systems
  • Expertise with multiple observability platforms and comparative knowledge
  • Background in APM (Application Performance Monitoring) and advanced monitoring techniques
  • Experience with machine learning-based anomaly detection and intelligent alerting
  • Knowledge of Open-Telemetry, or other sophisticated instrumentation technologies
  • Familiarity with chaos engineering and resilience testing
  • Security and compliance monitoring expertise (SIEM integration, audit logging)
  • Published articles, conference talks, or open-source contributions in observability domain
  • Relevant certifications (AWS, GCP, Azure, or vendor-specific observability certifications)
  • Experience in financial services or enterprise environments

More Info

Job Type:
Industry:
Function:
Employment Type:

Key Skills

logging strategies

audit logging

end-to-end transaction monitoring

synthetic monitoring

Open-Telemetry

distributed tracing

monitoring architectures

Alertsite

alerting frameworks

chaos engineering

SIEM integration