

Search by job, company or skills

Lead - Site Reliability Engineer
Location: Bengaluru (Hybrid)
About Setu
At Setu, we're building India's financial infrastructure.
Every day, millions of users rely on digital payments without realizing the complex systems that power them. Our mission is to simplify this infrastructure by building reliable APIs and platforms that enable businesses to launch financial products faster.
As part of Pine Labs, Setu powers critical payment infrastructure including UPI, Account Aggregator, BBPS, and Credit. We're looking for engineers who believe Every Day is Game Day and enjoy solving large-scale reliability and operational challenges.
About the Role
We're looking for SRE Lead to own the production reliability and operational excellence of our BBPS platform.
This role sits at the intersection of Engineering, SRE, DevOps and Customer Operations. You'll build an automation-first Production Engineering function that improves platform reliability, streamlines operational processes, and enables Core Engineering teams to focus on building products instead of repetitive operational work.
If you enjoy building systems, automating everything possible, improving reliability, and leading high-impact engineering initiatives, we'd love to talk.
What You'll Do
What We're Looking For
Nice to Have
You'll Thrive Here If You
Why Join Setu
At Setu, you'll work on infrastructure that powers millions of financial transactions every day. You'll solve hard engineering problems, collaborate with talented teams, and build systems that have a real impact on India's digital economy.
If building reliable, scalable financial infrastructure excites you, we'd love to hear from you.
Job ID: 151654175
Skills:
Agile Scrum, Puppet, Ansible, Incident Management, Cloudformation, Terraform, cloud cost optimization, infrastructure-as-code, cost allocation, reliability engineering, Production Support, tagging, Linux systems knowledge, Budget Tracking, AWS billing analysis, Forecasting, FinOps
Skills:
Yaml, Oracle Database, Bash, Json, Incident Response, Git, Bitbucket, Terraform, Linux, Rest Apis, Python, Performance Optimization, Intelligent automation, Root Cause Analysis, Agile engineering practices, Capacity Planning, Operational analytics, AI-assisted development, Site Reliability Engineering, Observability, Monitoring
Skills:
Elk, Prometheus, Grafana, Jenkins, Terraform, Python, Kubernetes, Infrastructure as Code, GKE, Go, AKS, EKS, GitLab CI, GitHub Actions, OpenTelemetry, ArgoCD
Skills:
Linux, Elasticsearch, Prometheus, Python, Kubernetes, Go
Skills:
PostgreSQL, Prometheus, Kafka, Grafana, Terraform, Gitlab, Redis, New Relic, Load Balancers, Jenkins, Iam, Vault, Helm, Kubernetes, Cloud Logging, Cloud NAT, GitOps, Cloud DNS, Cloud SQL, Cloud Armor, Cloud Monitoring, Shared VPC, Private Service Connect, Artifact Registry, Secret Manager, ArgoCD