Search by job, company or skills

Senior Site Reliability Engineer

  • Posted 14 hours ago
  • Be among the first 10 applicants

Job Description

Lead - Site Reliability Engineer

Location: Bengaluru (Hybrid)

About Setu

At Setu, we're building India's financial infrastructure.

Every day, millions of users rely on digital payments without realizing the complex systems that power them. Our mission is to simplify this infrastructure by building reliable APIs and platforms that enable businesses to launch financial products faster.

As part of Pine Labs, Setu powers critical payment infrastructure including UPI, Account Aggregator, BBPS, and Credit. We're looking for engineers who believe Every Day is Game Day and enjoy solving large-scale reliability and operational challenges.

About the Role

We're looking for SRE Lead to own the production reliability and operational excellence of our BBPS platform.

This role sits at the intersection of Engineering, SRE, DevOps and Customer Operations. You'll build an automation-first Production Engineering function that improves platform reliability, streamlines operational processes, and enables Core Engineering teams to focus on building products instead of repetitive operational work.

If you enjoy building systems, automating everything possible, improving reliability, and leading high-impact engineering initiatives, we'd love to talk.

What You'll Do

  • Build and lead the Production Engineering function for BBPS.
  • Own production health across availability, latency, success rates and operational excellence.
  • Design and improve incident management, on-call processes and escalation models.
  • Drive alert quality by eliminating noise and improving monitoring coverage.
  • Own production readiness for deployments, onboarding, disaster recovery and operational launches.
  • Improve incident RCA quality and ensure preventive actions are implemented.
  • Build automation for repetitive operational tasks, evidence collection, deployment validation and production workflows.
  • Partner closely with Core Engineering, DevOps, Solution Engineering, Product and InfoSec teams.
  • Mentor engineers and establish engineering best practices, runbooks and operational standards.

What We're Looking For

  • 6–10+ years of experience in Production Engineering, SRE, Platform Engineering or DevOps.
  • Strong experience managing production systems at scale.
  • Solid understanding of Linux, networking and distributed systems.
  • Experience with cloud platforms (AWS preferred).
  • Hands-on experience with Kubernetes and containerized environments.
  • Strong monitoring and observability experience using tools like Grafana, Prometheus and logging platforms.
  • Experience driving incident management, RCA and production support processes.
  • Strong scripting or programming skills (Python, Go, Bash or similar).
  • Passion for automation and eliminating repetitive operational work.
  • Excellent stakeholder management and cross-functional collaboration skills.

Nice to Have

  • Experience working with payment systems or fintech platforms.
  • Knowledge of BBPS, UPI or NPCI ecosystem.
  • Experience building SRE or Production Engineering teams.
  • Familiarity with compliance, DR exercises and production readiness reviews.

You'll Thrive Here If You

  • Take ownership of production systems end-to-end.
  • Enjoy solving operational problems through automation.
  • Prefer preventing incidents over reacting to them.
  • Can balance engineering quality with execution speed.
  • Like working across teams to improve reliability at scale.

Why Join Setu

At Setu, you'll work on infrastructure that powers millions of financial transactions every day. You'll solve hard engineering problems, collaborate with talented teams, and build systems that have a real impact on India's digital economy.

If building reliable, scalable financial infrastructure excites you, we'd love to hear from you.

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 151654175

Similar Jobs

Bengaluru, India

Skills:

Agile ScrumPuppetAnsibleIncident ManagementCloudformationTerraformcloud cost optimizationinfrastructure-as-codecost allocationreliability engineeringProduction SupporttaggingLinux systems knowledgeBudget TrackingAWS billing analysisForecastingFinOps

Bengaluru, India

Skills:

YamlOracle DatabaseBashJsonIncident ResponseGitBitbucketTerraformLinuxRest ApisPythonPerformance OptimizationIntelligent automationRoot Cause AnalysisAgile engineering practicesCapacity PlanningOperational analyticsAI-assisted developmentSite Reliability EngineeringObservabilityMonitoring

Bengaluru, India

Skills:

ElkPrometheusGrafanaJenkinsTerraformPythonKubernetesInfrastructure as CodeGKEGoAKSEKSGitLab CIGitHub ActionsOpenTelemetryArgoCD

Bengaluru

Skills:

LinuxElasticsearchPrometheusPythonKubernetesGo

Bengaluru, India

Skills:

PostgreSQLPrometheusKafkaGrafanaTerraformGitlabRedisNew RelicLoad BalancersJenkinsIamVaultHelmKubernetesCloud LoggingCloud NATGitOpsCloud DNSCloud SQLCloud ArmorCloud MonitoringShared VPCPrivate Service ConnectArtifact RegistrySecret ManagerArgoCD