

Search by job, company or skills

This job is no longer accepting applications
Job Title: Lead Site Reliability Engineer (SRE)
Location: Bangalore, Hyderabad & Chennai (Hybrid)
Experience: 8–14 Years
Job Summary
We are looking for a Lead Site Reliability Engineer (SRE) to lead a team of 5–6 engineers in a 24x7 production support environment. The ideal candidate should have strong experience in Incident Management, Linux, Kubernetes, Cloud (AWS/Azure/GCP), DevOps, Automation, and Production Operations with the ability to drive reliability, automation, and continuous improvement.
Key Responsibilities
Mandatory Skills
Preferred
Looking for professionals who can thrive in high-pressure production environments, lead critical incidents, and drive automation to improve system reliability.
Job ID: 151081043
Skills:
Java, Golang, Google Cloud Platform, Kafka, Spark, Azure, Kubernetes, Python, AWS, Airflow, Flink, dbt, MLFlow, Large Language Models
Skills:
Devops, Site Reliability Engineering
Skills:
Ml, Java, Spring Boot, Jenkins, Docker, Terraform, ECS, Gitlab, Kubernetes, Python, AWS, Open Telemetry, Ai, Site Reliability Engineering
Skills:
Ml, Java, Spring Boot, Jenkins, Docker, Terraform, ECS, Gitlab, Kubernetes, Python, AWS, Open Telemetry, Ai, Site Reliability Engineering
Skills:
Tcp, UDP, Dns, Rtp, Load Testing, Gcp, Load Balancing, Tls, Python, incident communication, automatic rollback, NATS-class buses, capacity modeling, Go, canary analysis, WebSocket fleets, AI-aware reliability, multi-cloud literacy, SIP, streaming pipelines, chaos engineering