

Search by job, company or skills
Senior DevOps Engineer
Next Generation Monitoring (NGM) Platform - Vertiv IT Systems | Pune, Maharashtra
Experience: 8-12 Years . Band: Senior Engineer (IC) . Department: Software Engineering - IT Systems . Reports To: Engineering Manager
POSITION SUMMARY
We are seeking an experienced Senior DevOps Engineer to own the engineering delivery infrastructure - CI/CD pipelines, container orchestration, multi-deployment packaging, release management, security automation, and platform observability. The engineer will support on-premises, cloud-native (AWS/GCP/Azure), hybrid edge-cloud, and air-gapped deployment targets across a multi-year, multi-release product roadmap.
KEY RESPONSIBILITIES
CI/CD Pipeline Engineering
. Design, build, and optimize GitLab CI/CD pipelines for all microservices with independent deployment cycles
. Implement multi-stage pipelines: build → test → security scan → package → deploy across all environments
. Integrate automated unit, integration, and system-level tests manage all pipeline configuration as code
. Configure branch-based promotion strategies (dev → staging → production)
Container Orchestration & Infrastructure
. Manage containerized deployments via Docker and Kubernetes on Linux and Windows Server
. Provision and manage Kubernetes clusters (dev, staging, production) with horizontal pod autoscaling
. Configure service mesh (Istio/Linkerd) for inter-service mTLS manage container image registry lifecycle
. Implement Infrastructure as Code using Terraform and/or Ansible for cloud and on-premises resources
Multi-Deployment Operations
. Build on-premises Docker/Kubernetes deployment automation for Linux and Windows Server targets
. Develop cloud-native automation for AWS, GCP, and Azure (EKS/GKE/AKS) including multi-region config
. Engineer hybrid edge-cloud pipelines package air-gapped artefact bundles for offline/classified environments
. Implement zero-downtime rolling updates, blue-green deployments, and automated rollback strategies
Release Management
. Orchestrate software releases across the multi-release product roadmap coordinate cross-functional readiness
. Manage versioning, release tagging, changelog generation, and artefact storage across all microservices
. Produce deployment runbooks, release notes, and rollback procedures for each release milestone
Security & Compliance Automation
. Integrate SAST, DAST, container image scanning (Trivy/Snyk/Clair), and dependency vulnerability checks
. Automate SBOM generation, open-source license compliance, and software supply chain tracking
. Configure secrets management (HashiCorp Vault), TLS certificate automation, and PKI for microservices
. Implement pipeline controls for SOC 2, EU Cyber Resilience Act, and FedRAMP readiness
Monitoring, Observability & Reliability
. Deploy centralized logging (ELK/Loki), metrics (Prometheus/Grafana), and distributed tracing (Jaeger)
. Configure Kubernetes cluster monitoring, pod health dashboards, and alerting/on-call integration
. Drive proactive reliability engineering to meet high-availability platform SLA targets
Automation & Cross-Functional Collaboration
. Create Bash/Python automation for infrastructure operations, environment provisioning, and self-healing
. Interface with software architects, developers, QA, and product management on delivery and planning
. Guide development teams on DevOps best practices, containerization patterns, and pipeline design
REQUIRED TECHNICAL SKILLS
Infrastructure & Platform: Linux (Ubuntu, RHEL, CentOS) production admin (5+ yrs) . Docker & Kubernetes (Helm, RBAC, autoscaling) . Service mesh: Istio/Linkerd . Windows Server 2019+ for on-premises targets
CI/CD & Version Control: GitLab CI/CD administration & pipeline YAML (5+ yrs) . Multi-service microservices pipelines . Artefact management: GitLab, ECR, GCR, ACR . GitOps with Helm chart management
Cloud & Multi-Deployment: AWS, GCP, or Azure (multi-cloud preferred) . EKS/GKE/AKS auto-scaling & multi-region . Hybrid edge-cloud . Air-gapped offline artefact packaging . Terraform & Ansible (5+ yrs)
Development & Automation: Advanced Bash & Python scripting . YAML/JSON for Kubernetes manifests & Helm . REST API automation . Go (Golang) familiarity for microservices platform context
Security & Compliance: Container scanning: Trivy, Snyk, or Clair . SAST/DAST pipeline integration . SBOM: Syft/CycloneDX . HashiCorp Vault . TLS/PKI automation . SOC 2, EU CRA, FedRAMP pipeline controls
Monitoring & Observability: ELK/EFK, Loki, or Splunk . Prometheus & Grafana (production-grade) . Distributed tracing: Jaeger/Zipkin . Kubernetes cluster monitoring . SLA reliability metrics
GOOD-TO-HAVE SKILLS
. Cloud certifications: AWS DevOps Engineer, Google Professional DevOps, or Azure DevOps Expert
. CKA/CKAD, Kubernetes operator development, and service mesh deep expertise
. GitOps tooling: ArgoCD or Flux for continuous delivery
. SRE practices: error budget management, chaos engineering (LitmusChaos, Chaos Monkey)
. Time-series DB operations: InfluxDB or TimescaleDB backup, scaling, schema migration
. Message broker management: Kafka or NATS cluster operations and monitoring
. Edge node provisioning, remote management, and telemetry collection from distributed agents
. Internal Developer Platform (IDP) tooling and developer self-service workflows
REQUIRED QUALIFICATIONS
. Bachelor's Degree in Computer Science, Information Technology, Electrical Engineering, or related field
. Minimum 8 years of experience in DevOps, Site Reliability Engineering, or Infrastructure Engineering
. 5+ years of Linux system administration in production environments (Ubuntu, RHEL, CentOS)
. 5+ years of Docker and Kubernetes container orchestration in production
. 5+ years of CI/CD pipeline management (GitLab CI preferred)
. 5+ years of cloud platform deployments: AWS, GCP, or Azure
. Proven track record delivering end-to-end CI/CD pipelines for microservices architectures
. Strong background in Infrastructure as Code (Terraform, Ansible, or equivalent)
. Experience integrating security and compliance tooling into DevOps workflows
SOFT SKILLS & EDUCATION
. Strong problem-solving and root-cause analysis in complex distributed systems
. Clear written and verbal communication for cross-functional engineering collaboration
. Proactive ownership mindset with a strong bias toward automation over manual operations
. Ability to manage competing priorities in a fast-paced agile product development environment
. Team collaboration and informal DevOps culture leadership comfort with ambiguity
Education: B.E. / B.Tech / M.E. / M.Tech in Computer Science, IT, or related field. Cloud/DevOps certifications (CKA/CKAD, AWS/GCP/Azure) are a plus.
Job ID: 153931077