Search by job, company or skills

Site Reliability Engineer

Site Reliability Engineer

zorba ai
8-10 Years
Not Disclosed
Early Applicant
  • Posted 6 days ago
  • Be among the first 10 applicants

Job Description

Job Description

  • mandatory

SN

Required Information

Details

1

Role**

Site Reliability Engineer

2

Required Technical Skill Set**

Azure/GCP, .NET, DevOps Practices, ITIL, ReactJS, React Native, NodeJS, Terraform/Ansible, Python, C#, Kubernetes, Splunk, Dynatrace.

3

No of Requirements**

15

4

Desired Experience Range**

8years – 10 years

5

Location of Requirement

Chennai/Hyderabad/Kochi/Banglore

Desired Competencies (Technical/Behavioral Competency)

Must-Have**

  • Work in a cross functional team working with Reliability as Expertise in a product or a product area.
  • Apply Reliability engineering practices with support from SRE governance teams.
  • 5+ years of experience in Site Reliability Engineering, maintenance & operations and/or development.
  • Strong working experience eCommerce.
  • Strong working experience in DevOps practices (automated testing, CI/CD etc.).
  • Experience within solutions architecture and how to fast pinpoint causes of issues.
  • Experience from working with API-based frameworks (e.g., Commerce tools or Fabric is ideal).
  • Experience in building CI/CD workflows using GitHub Actions.
  • Experience of maintaining/supporting and/or developing desktop and mobile applications.
  • Experience working on cloud-based infrastructure e.g., Azure and GCP.
  • Experience in provisioning Infra resources leveraging Infra as Code (Terraform / Ansible).
  • Experience from ITIL support processes and ITSM tools (e.g., ServiceNow) in a microservices context.
  • Experience in monitoring tools (Splunk, Grafana etc.).
  • Experience working through SRE Metrics such as SLI, SLO and Error Budget.
  • Experience with managed cloud Kubernetes services (e.g. AKS, GKE).
  • Ensure delivery quality and supply KPI reporting.
  • Collaborate closely within product teams to ensure predictable operations and minimal disruptions to Production.
  • Collaborate closely within your Capability, share best practices as well as discuss and improve on operations ways of working.
  • Technical analysis, troubleshooting of complex issues/Incidents in production.
  • Improve monitoring performance by focusing on preventive measures.
  • Product Improvements (code & log analysis).
  • Continuous improvement on proactive monitoring, housekeeping automation to proactively detect and avoid incidents.
  • Automate processes impacting development and production leveraging tools and building scripted solutions.
  • Participate in On-Call technical support to resolve business critical incidents.

Good to have

  • Familiarity with common tech stacks in Headless Ecommerce is a nice to have.
  • Knowledge of design principles and fundamentals of solutions architecture is a plus.
  • Understanding of performance engineering (Application Reliability).
  • Knowledge of multiple front-end languages and libraries (ReactJS, React Native, NodeJS).
  • Knowledge of Azure DevOps and/or other cloud environments is nice to have.
  • A passion for problem solving with strong analytical capabilities.
  • Stay current on technical trends to suggest innovative tools and approaches to interesting problems.
  • Knowledge on at least one of Python, Ruby, Java, C#, Go at an intermediate level.

SN

Responsibility of / Expectations from the Role

1

Should be able to analyze the critical issue and resolve in time

2

Willingness to accept development challenges and work towards resolving them

3

Should be a good team player and should be able to lead the team

4

Should be willing to work in shifts

5

Should be focused in on-time delivery

Type

Details of The Role (For Candidate Briefing)

Reporting To Which Role

Site Reliability Engineer

Size of the Team, if any Reporting to this Role

Not Applicable

On-site Opportunity

No Visibility

Unique Selling Proposition (USP) of The Role

Opportunity to work in multiple technologies & directly with Customer.

Details of The Project (A short Briefing on the Project may be attached with this document for candidate- briefing). It may be shared with external stakeholders like job-agencies etc.

Providing technical leadership to teams and integrate technical expertise and business understanding to collaborate with Customer involved in Retail Business in all over the world

Skills: reliability,devops,azure,ansible

More Info

Job Type:
Industry:
Employment Type:

About Company

Similar Jobs

5-8 yrs
Hyderabad, India
Skills:
.NET, Java, Prometheus, Nodejs, Azure Log Analytics, Grafana, JIRA, Datadog, Gcp, Docker, Splunk, Azure, Kubernetes, Python, AWS, PagerDuty
8-10 yrs
Hyderabad, India
Skills:
Context Protocol (MCP, Real User Monitoring (RUM, Itsm, Dynatrace OneAgent, Distributed Tracing, Synthetic Monitoring, Davis AI, AI-powered Self-Service Agents, Smartscape, Davis AI Model, Agile practices, Servers Claude
6-9 yrs
Hyderabad, India
Skills:
S3, RDS, Bash, Datadog, Jenkins, Lambda, Ec2, Terraform, Ansible, Python, Kubernetes, AWS
10-12 yrs
Hyderabad, India
Skills:
Terraform, Ansible, PostgreSQL, Prometheus, Grafana, Helm, Kubernetes, AWS, GitHub Actions
9-12 yrs
Hyderabad, India
Skills:
bedrock , S3, Prometheus, Vpc, Grafana, Datadog, Lambda, Ec2, Terraform, Docker, Python, RDS, Kms, Cloudwatch, Splunk, Kubernetes, Infrastructure as Code, Secrets Manager, GitOps, AWS architecture, Sagemaker, PrivateLink, OpenTelemetry, CloudTrail, EKS