Search by job, company or skills

Site Reliability Engineer

Early Applicant
  • Posted a month ago
  • Be among the first 10 applicants

Job Description

HighRadius intends to build new products as a part of new initiatives entrepreneurship program. These new products are being built in a startup ecosystem.

We are seeking a highly skilled and experienced Senior Cloud Engineer to join our dynamic team. The ideal candidate will be instrumental in designing, implementing, and managing our cloud infrastructure, with a strong focus on automation, scalability, and reliability. You will play a key role in evolving our DevOps practices, ensuring seamless integration of development and operations, and contributing to the overall success of our cloud-native initiatives.

Responsibilities:

  • Design, implement, and manage scalable, secure, and highly available cloud infrastructure primarily on AWS.
  • Develop and maintain Infrastructure as Code (IaC) using Terraform to automate provisioning and configuration of cloud resources.
  • Implement and manage CI/CD pipelines using Jenkins and GitHub/GitLab to automate software delivery, including build, test, and deployment processes.
  • Orchestrate and manage containerized applications using Docker, Kubernetes, and Helm.
  • Establish and maintain robust monitoring and logging solutions using Prometheus, Grafana, and either Datadog, Splunk, or Kibana to ensure system health and performance.
  • Develop and maintain automation scripts using Python, Shell, or GoLang to streamline operational tasks and improve efficiency.
  • Champion DevOps best practices, including GitOps, security integration throughout the SDLC, and advanced deployment strategies like Blue/Green deployments.
  • Collaborate closely with development, operations, and security teams to ensure the smooth functioning and continuous improvement of our cloud platforms.
  • Troubleshoot and resolve complex infrastructure and application issues in a timely manner.
  • Participate in on-call rotations as needed to support critical systems.
  • Mentor junior engineers and contribute to the growth of the team's technical capabilities.
  • Stay up-to-date with the latest cloud technologies, trends, and best practices.

Required Skills and Experience:

  • We are looking for Minimum 7+ years of experience in SRE (Site Reliability)

Cloud: Extensive hands-on experience with Amazon Web Services (AWS) including:

  • Compute: EC2
  • Storage: S3
  • Identity & Access Management: IAM
  • Container Orchestration: EKS
  • Infrastructure as Code: Terraform

CI/CD: Proven experience in designing and implementing robust CI/CD pipelines with:

  • Jenkins
  • GitHub/GitLab
  • Maven (or similar build tools)

Containerization: Deep understanding and practical experience with:

  • Docker
  • Kubernetes
  • Helm

Infrastructure as Code (IaC):

  • Terraform (mandatory)
  • Ansible (highly desirable)

Monitoring & Logging: Experience in setting up and managing monitoring and logging solutions using at least two of the following:

  • Prometheus
  • Grafana
  • Datadog/Splunk/Kibana

Scripting: Proficiency in at least two of the following scripting languages:

  • Python
  • Shell
  • GoLang

DevOps Practices: Strong understanding and practical application of:

  • GitOps methodologies
  • Integrating security into the DevOps pipeline (DevSecOps)
  • Advanced deployment strategies (e.g., Blue/Green Deployments)
  • Excellent problem-solving skills and the ability to diagnose and resolve complex technical issues.
  • Strong communication and interpersonal skills, with the ability to collaborate effectively with cross-functional teams.
  • Ability to work independently and as part of a team in a fast-paced, agile environment.

Preferred Qualifications:

  • AWS Certifications (e.g., AWS Certified Solutions Architect - Associate/Professional, AWS Certified DevOps Engineer - Professional).
  • Experience with other cloud providers (e.g., Azure, GCP) is a plus.
  • Familiarity with database technologies (SQL/NoSQL).
  • Experience with security best practices in cloud environments.

Education:

  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 127019687

Similar Jobs

Hyderabad, India

Skills:

JavaGolangGoogle Cloud PlatformKafkaSparkAzureKubernetesPythonAWSAirflowFlinkdbtMLFlowLarge Language Models

Hyderabad, India

Skills:

GitJavascriptDockerNode.jsRESTful API designMicroservices architecture

Hyderabad, India

Skills:

MlJavaSpring BootJenkinsDockerTerraformECSGitlabKubernetesPythonAWSOpen TelemetryAiSite Reliability Engineering

Hyderabad

Skills:

MlJavaSpring BootJenkinsDockerTerraformECSGitlabKubernetesPythonAWSOpen TelemetryAiSite Reliability Engineering

Hyderabad, India

Skills:

TcpUDPDnsRtpLoad TestingGcpLoad BalancingTlsPythonincident communicationautomatic rollbackNATS-class busescapacity modelingGocanary analysisWebSocket fleetsAI-aware reliabilitymulti-cloud literacySIPstreaming pipelineschaos engineering