Site Reliability Engineer - Enterprise SaaS product - Hyderabad
CareerXperts Consulting- Posted 7 hours ago
- Be among the first 10 applicants
Job Description
We're looking for a Senior Platform Engineer with a software engineering background to join our globally dispersed Site Reliability and Platform team, operating a multi-region cloud platform in a follow the sun model.
You bring strong software engineering fundamentals, DevOps/infrastructure chops, and cloud platform expertise — applying that experience to improve system performance, reliability, and automate away manual work.
What You'll Own:
Technical leadership and mentoring — knowledge sharing, pair programming, code reviews, solution design
Platform reliability improvements — mitigation strategies and operational playbooks
Monitoring, alerting, and logging systems for incident detection and response
RCAs and blameless post-mortems; on-call rotation participation
Cloud infrastructure scalability and performance testing
Infrastructure as Code — automated provisioning, configuration, and management
Database monitoring and tuning for high availability and performance
Defining SLIs, SLOs, and SLAs with cross-functional teams
What you bring:
3+ years in Platform Engineering, SRE, or Software Engineering
1+ years working on a SaaS platform
Strong software engineering background (any modern language)
Production experience with Kubernetes
Proficiency with Azure, AWS, or GCP
IaC tooling — Terraform (preferred), Ansible, or CloudFormation
Scripting/automation — Bash, PowerShell, or Python
Monitoring/observability tools — DataDog, Prometheus, Grafana, or similar
Track record maintaining highly-available, performant production environments
Bonus: CI/CD tooling (Azure DevOps/GitHub Actions, Octopus Deploy), cloud/DevOps certifications, database performance tuning experience.
If you love turning fragile systems into reliable ones — let's talk. Write to [Confidential Information]
