Site Reliability Engineer - Chennai-Virtual drive -5th Sep-Saturday
HCL TechBee- Posted 16 days ago
- Be among the first 20 applicants
Job Description
Skill Requirements
Must Have Skills:
- experience in Production Operations, SRE, Observability, Application Support, or Operations Engineering roles.
- Experience with observability, logging, alert management, alert correlation, and alert quality improvement practices using Splunk ITSI, Splunk Observability Cloud & Splunk Enterprise logging.
- Experience implementing and supporting SLI, SLO, Error Budget tracking, service health monitoring, reliability reporting, and continuous improvement initiatives.
- Experience with Incident Management, Problem Management, Troubleshooting, Change Management, Root Cause Analysis (RCA), and Blameless Postmortem practices.
- Experience supporting business applications running on VM and container platforms.
- Experience with cloud platforms such as Azure, AWS, and GCP.
- Experience with automation technologies to reduce the operational toil (Ansible, Python, RPA and other automation platforms)
- Experience with ITSM processes and tools such as ServiceNow.
- Strong analytical, troubleshooting, communication, and collaboration skills.
Please fill the below form to apply for this job.
SRE -Drive- 5th Sep 26 – Fill out form
More Info
Key Skills
Splunk ITSI
Splunk Observability Cloud
container platforms
reliability reporting
SLI
Splunk Enterprise logging
Error Budget tracking
SLO
Blameless Postmortem
service health monitoring
