Search by job, company or skills

Site Reliability Engineer

Site Reliability Engineer

Tata Consultancy Services
8-10 Years
Not Disclosed
  • Posted 8 hours ago
  • Be among the first 10 applicants

Job Description

Site Reliability Engineer (SRE)

Location: Bangalore, India

Experience: 8 to 10 Years

We are looking for an experienced Site Reliability Engineer (SRE) with strong expertise in cloud infrastructure, automation, observability, and platform reliability. The ideal candidate will be responsible for ensuring the availability, scalability, performance, and security of mission-critical applications and infrastructure hosted on Microsoft Azure.

Key Responsibilities

  • Design, implement, and maintain highly available and scalable cloud infrastructure on Microsoft Azure.
  • Define and drive SRE best practices, including reliability engineering, incident management, observability, and automation.
  • Manage and optimize Kubernetes clusters for containerized workloads.
  • Build and maintain Infrastructure as Code (IaC) solutions using Terraform.
  • Implement CI/CD pipelines using Azure DevOps and support deployment automation.
  • Monitor system health, application performance, and infrastructure using Splunk and other monitoring tools.
  • Lead root cause analysis (RCA), incident response, and problem management activities.
  • Collaborate with development, platform, and operations teams to improve system reliability and release quality.
  • Establish SLIs, SLOs, and error budgets to enhance service reliability.
  • Automate operational tasks and improve platform efficiency.

Required Skills (Must Have)

  • Strong experience as a Site Reliability Engineer (SRE).
  • Hands-on expertise in Microsoft Azure Cloud.
  • Extensive experience with Azure DevOps.
  • Strong knowledge of Kubernetes administration and troubleshooting.
  • Experience implementing IaC using Terraform.
  • Proficiency in Splunk for monitoring, logging, and observability.
  • Strong understanding of CI/CD, automation, system reliability, and performance optimization.
  • Experience with Linux systems, scripting, and production support environments.

Good to Have

  • Experience with GitHub Actions.
  • Knowledge of DevSecOps practices.
  • Exposure to container security and cloud-native observability tools.
  • Experience with multi-cloud environments.

More Info

Job Type:
Industry:
Employment Type:

Key Skills

Similar Jobs

8-12 yrs
Bengaluru, India
Skills:
Terraform, PostgreSQL, Prometheus, Dynatrace, Grafana, OpenTelemetry, Istio, AlloyDB
10-15 yrs
Bengaluru, India
Skills:
VMware, Kvm, Linux, Ansible, Bash, Kubernetes, Python
8-12 yrs
Bengaluru, India
Skills:
Saml, Prometheus, Kafka, Spring Boot, Datadog, Docker, Terraform, Teamcity, Python, AWS, Oauth, Java, RDS, Sso, Jenkins, Cloudwatch, Bitbucket, Sqs, Helm, Kubernetes, Go, Aurora, GitHub Actions, OpenTelemetry, FluxCD
5-8 yrs
Bengaluru, India
Skills:
Unix, Cloudformation, Prometheus, Bash, Grafana, Cloudwatch, Docker, Terraform, Linux, Ansible, Splunk, Kubernetes, Python, AWS, Open Telemetry, Go, EKS
6-13 yrs
Bengaluru, India
Skills:
Linux Administration, Incident Management, Dynatrace, Splunk, Azure, Java Application Support, Observability