Search by job, company or skills

Platform Software Engineer

Platform Software Engineer

People Gamut HR Solutions
6-9 Years
Not Disclosed
Quick Apply
  • Posted 19 days ago
  • Over 100 applicants have applied

Job Description

Job Description

We are looking for a Senior Site Reliability Engineer who is passionate about building and running a high-performance cloud platform and enabling best-in-class site reliability and operations practices. This role will support our operations globally. The candidate will drive the development of modern, cloud-native SRE processes and the management and operations for our multi-tenant, microservices-based cloud platform. The platform has multiple instances deployed across the globe.

This role involves working closely with cross-functional teams to integrate reliability and security into our systems, ensuring they meet standards. The ideal candidate will have extensive experience in both software engineering and systems administration, with a strong understanding of SRE concepts, requirements and security practices.

Infrastructure Management:

•            Oversee the design, deployment, and management of scalable and secure cloud infrastructure.

•            Drive automation of infrastructure provisioning, configuration, and management using Infrastructure as Code (IaC) tools.

Monitoring and Performance:

•            Develop and maintain comprehensive monitoring, logging, and alerting systems to ensure high availability and performance.

•            Contribute to performance tuning and optimization for applications and infrastructure.

Security and Compliance:

•            Ensure implementation and maintenance of security controls and best practices to achieve compliance with standards and certifications.

•            Conduct and oversee regular security assessments, vulnerability scans, and penetration testing.

•            Collaborate with the compliance team to prepare for and respond to audits.

Incident Management:

•            Manage incident management efforts, ensuring rapid resolution and thorough root cause analysis.

•            Develop and implement strategies for improving incident response and minimizing downtime.

 

Collaboration and Communication:

•            Work closely with development, operations, and security teams to integrate reliability and security into the software development lifecycle.

•            Communicate effectively with stakeholders, providing regular updates on system performance, reliability, and compliance status.

Qualifications

•            Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent experience).

•            6+ years of experience in site reliability engineering, DevOps, or a related role.

•            Proficiency in cloud platforms (AWS, Azure, GCP) and cloud-native services.

•            Strong scripting and programming skills (Python, Bash, Go, or similar).

•            Experience with Infrastructure as Code (IaC) tools such as Terraform, CrossPlane, CloudFormation, or Ansible.

•            Knowledge of containerization and orchestration (Docker, Kubernetes).

•            Familiarity with CI/CD pipelines and tools (Jenkins, GitLab, GitHub, etc.).

•            In-depth knowledge of standards (ISO, SOC2...) requirements and best practices.

•            Experience with security tools and practices (SIEM, IDS/IPS, firewalls).

•            Understanding of network security, encryption, and secure software development practices.

•            Ability to collaborate with and foster effective communication with global and multicultural engineering teams in EU and US timezones.

•            Ability to report timely and effectively to the upper engineering management.

Additional Information

We are the pioneers and trailblazers of a global IT Market Category (DEX) that is shaping the future of how the world works, giving our customers IT Teams total digital visibility across their enterprise. Our innovative solutions integrate real-time analytics, automation, and employee feedback across all endpoints. This enables our IT teams to solve complex technical challenges, create ever more productive workplaces, and deliver happy, satisfied employees in the digital workplace.

With over 1000 employees across 5 continents, we operate as One Team, connecting, collaborating and innovating to continuously grow. Our commitment to diversity, inclusion, and equity is second to none. We currently have over 75 nationalities working with us, from all cultures and backgrounds, speaking many different languages.

You can reach out to me - 8050030856 or share your resumes to [Confidential Information] to apply.

More Info

Job Type:
Function:
Employment Type:

Key Skills

User Avatar

About Recruiter

Amrith Parameshwar

0 Active Jobs

Similar Jobs

10-12 yrs
Bengaluru, India
Skills:
Algorithms, concurrency, Software Design, Distributed Systems, Cloud Infrastructure, data structures, networking software, Systems Software, large-scale distributed systems, highly available production-grade distributed services, CI CD pipelines, release automation, network routing protocols, deployment workflows
5-7 yrs
Bengaluru, India
Skills:
Vpc, AWS, Kubernetes, CloudFront, Terraform, Api Gateway, EKS, OpenTofu, GitOps
5-7 yrs
Bengaluru, India
Skills:
Java, Golang, PostgreSQL, Node.js, HTML, Angular, C Sharp, React, Docker, MySQL, Integration Testing, Kubernetes, AWS
10-12 yrs
Bengaluru, India
Skills:
.NET, Bdd, Nodejs, Angular, Nosql, Tensorflow, React, Pytorch, Docker, Terraform, Selenium, Python, Java, Rust, Big Data, Sql, Gherkin, MLops, Databricks, Cucumber, Kubernetes, LangFuse, LangSmith, LangGraph, OpenTelemetry, LangChain, Playwright, LLMOps
6-8 yrs
Bengaluru, India
Skills:
Linux Os, Docker, Terraform, Prometheus, Bash, Python, Kubernetes, AWS, Microservices, ArgoCD, UNIX/Linux based systems, CI/CD, Container technologies, PagerDuty