

Search by job, company or skills

Required Skills
· 8–10 years of overall IT experience.
· Strong expertise in Linux administration and troubleshooting.
· Hands-on experience with at least one of the following:
· Strong debugging and root cause analysis skills across application and infrastructure layers.
· Experience supporting and managing on-premises production environments.
· Excellent operational discipline with experience in incident management, change management, and production support.
· Good understanding of system monitoring, logging, and performance tuning.
· Strong scripting skills (Shell/Python) are an added advantage.
· Automation and Zero touch mindset for next level GenAI adoption
Preferred Qualifications
· Experience with automation and configuration management tools.
· Knowledge of container platforms (Docker/Kubernetes) is a plus.
· Familiarity with CI/CD pipelines and DevOps practices.
· Excellent communication and stakeholder management skills.
Job ID: 151989167
Skills:
automation, Gpu, Troubleshooting, performance optimization, Site Reliability Engineering
Skills:
Google Cloud Platform, Linux, Dns, Azure, Tls, Kubernetes, Python, AWS, Cilium, Go, Istio
Skills:
Datadog, AWS, Kubernetes, Python, Terraform, Grafana, Shell, ClickHouse, VictoriaMetrics, Go, GitOps
Skills:
Networking, Splunk, automation, Datadog, AWS, cloud, Python, Azure, Gcp, Terraform, observability, IaC, Alibaba Cloud, Infrastructure-as-Code, Zero Trust, CI CD, Entra ID, OIDC, security tooling
Skills:
Grafana, Bitbucket, Apigee, Github, Kafka, Newrelic, Prometheus, Kubernetes, Bash, Python, Docker, Terraform, Jenkins, Git, Microsoft Azure, Elk Stack, Azure DevOps, Azure Monitor, GitLab CI, Go