Platform Engineer Pune
- Posted 13 hours ago
- Be among the first 10 applicants
Job Description
Experience: 5–10 Years
Location: Pune
Notice Period: Immediate to 45 Days(685)
Shift: 24×7 Operations
Qualification: Bachelor's degree in CS, IT, Engineering or related field
Role Overview
We are hiring a Platform Engineer to support a 24×7 AI platform operations environment. The role involves platform monitoring, incident troubleshooting, recovery, escalation and operational support across Kubernetes/OpenShift-based platforms.
Key Responsibilities
2 Technical Rounds + Client Round + HR
Skills: python scripting,kubernetes,red hat openshift,prometheus,runbooks,platform monitoring,grafana,incident troubleshooting,recovery,containers,networking,linux command-line,sops,strong communication skills
Location: Pune
Notice Period: Immediate to 45 Days(685)
Shift: 24×7 Operations
Qualification: Bachelor's degree in CS, IT, Engineering or related field
Role Overview
We are hiring a Platform Engineer to support a 24×7 AI platform operations environment. The role involves platform monitoring, incident troubleshooting, recovery, escalation and operational support across Kubernetes/OpenShift-based platforms.
Key Responsibilities
- Monitor platform/application dashboards, alerts and operational systems.
- Perform L1 incident detection, troubleshooting and recovery.
- Troubleshoot basic Linux, networking, DNS, ports and connectivity issues.
- Check Kubernetes/OpenShift nodes, pods, deployments and services.
- Review logs and collect diagnostics for escalation.
- Execute approved runbooks, SOPs and GitOps-based recovery procedures.
- Restart/redeploy workloads and verify service recovery.
- Create incident tickets, provide status updates and coordinate with L2/L3 teams.
- Maintain proper shift handovers and operational documentation.
- Platform Monitoring
- Incident Troubleshooting & Recovery
- Kubernetes & Red Hat OpenShift
- Experience with Linux command-line
- Experience with Networking: IP, DNS, ports & connectivity
- Experience with Containers knowledge
- Experience in IT Operations / Infrastructure / Cloud / Application Support
- Ability to follow SOPs, runbooks and technical procedures
- Strong troubleshooting and analytical skills
- Excellent communication & stakeholder management
- Willingness to work in a 24×7 shift environment
- Minimum 2 years of stability in an organization preferred
- Git / GitOps
- Grafana / Prometheus
- Bash / Python scripting
- ITIL / Incident Management
- Cloud / Data Centre infrastructure
- AI/ML or GPU infrastructure
- SRE / Platform Engineering exposure
- Strong communication skills are mandatory.
- Pune/local candidates preferred.
- Candidates should be comfortable working in a structured L1 operations and incident management environment.
2 Technical Rounds + Client Round + HR
Skills: python scripting,kubernetes,red hat openshift,prometheus,runbooks,platform monitoring,grafana,incident troubleshooting,recovery,containers,networking,linux command-line,sops,strong communication skills
More Info
Job Type:
Industry:
Function:
Employment Type:
Key Skills
GitOps
Runbooks
Platform Monitoring
Red Hat OpenShift
Incident Troubleshooting Recovery
Linux command-line
Containers


