Search by job, company or skills

Site Reliability Engineer - 2/3

Site Reliability Engineer - 2/3

Fynd
3-6 Years
Not Disclosed
Quick Apply
  • Posted a month ago
  • Over 100 applicants have applied

Job Description

What will you do at Fynd

  • Influence technical direction by evaluating change requests, participating in architectural discussions across teams to uphold best practices and decide on appropriate technologies.
  • Lead incident response and root cause analysis to rapidly resolve issues and implement preventive measures, ensuring we never fail for the same reason twice.
  • Identify any bottleneck in current processes and build or improve tools to support incident management.
  • Go on-call, respond to automated alerts, and execute playbooks.
  • Continuously monitor and fine-tune our infrastructure using industry-standard observability tools, ensuring high performance even under heavy load.
  • Conduct rigorous load tests for critical sales events and optimise system capacity to handle peak demand seamlessly.
  • Own availability and performance for key products. Be responsible for ensuring the product's architecture, changes, incident response, and technology choices support its target availability and performance levels.
  • Remove unnecessary noise from our signals to obtain a clearer understanding of our platform and enable more effective debugging.
  • Develop production tooling and services to improve our platform's resilience.

Minimum Qualification:

  • Bachelor's degree (B. E./B. Tech.) in Computer Science, or a related technical field, or equivalent practical experience.
  • 2+ years of experience in an SRE or DevOps role, preferably within the e-commerce sector.
  • 2+ years of experience in programming languages such as Go, Python, or JavaScript, coupled with a solid understanding of data structures and algorithms.
  • Experience with containerisation technologies such as Docker and Kubernetes.
  • Experience with cloud platforms like AWS, GCP, or Azure.
  • Experience with monitoring and alerting tools such as Grafana, Prometheus, Sentry, PagerDuty, New Relic, AWS CloudWatch, etc.
  • Proficiency in Unix/Linux shell environments.

Some specific Requirements:

  • 3+ years of experience in an SRE or DevOps role, preferably within the e-commerce sector.
  • 3+ years of experience managing production infrastructure. Prior experience leading or managing a team is a strong advantage.
  • Experience with message queues like Kafka or RabbitMQ and a strong understanding of event-driven architectures.
  • Experience with any orchestration and deployment tools such as Terraform, Pulumi, AWS CloudFormation, etc.
  • Hands-on experience with any configuration management systems like Ansible, Chef, Puppet, SaltStack, etc.
  • Understanding of load testing methodologies and tools such as Grafana k6, Gatling, Locust, Apache JMeter, etc.

More Info

Job Type:
Industry:
Function:
Employment Type:

Key Skills

About Company

Fynd is India's largest omnichannel ecosystem and multi-platform tech company. Headquartered in Mumbai and founded by Farooq Adam, Harsh Shah, and Sreeraman MG in 2012.

We have modernized retail strategies for more than 1000 brands & created a rich suite of tech products. Rooted in technology & innovation, we have products in applied machine learning, big data, gaming+crypto, image editing, and learning space.

Our constant innovation and expertise in technology has been noticed worldwide. Fynd made it to Fast Company's list of Top 10 most innovative Asia-Pacific companies of 2022.

We are a fast growing team of 1000+ fun, skilled and ambitious people. We explore the unexplored, innovate unafraid, and have the time of our life while we do. Be a part of the new. Join us.