Principal App and Infra Monitoring Lead
We are seeking App & Infra Monitoring Engineer to join our client Engineering team. She/He will be focusing on enabling monitoring and alerting strategies across multiple groups within client.
Qualifications:
- 3+ years in IT infrastructure, application support, or monitoring roles.
- Experience in incident response, troubleshooting, root cause analysis, and performance tuning.
- Monitoring Tools Expertise: Experience with tools like Datadog, Splunk, Dynatrace, AppDynamics, New Relic, Nagios, Prometheus, Grafana, etc.
- Infrastructure Monitoring: Familiarity with server, network, and cloud monitoring solutions across AWS, Azure, GCP, or on-premises data centers.
- Application Performance Monitoring (APM): Understanding of application logs, metrics, tracing, and alerting mechanisms.
- Scripting & Automation: Proficiency in Python, Shell, PowerShell, or Ansible for automating monitoring tasks.
- Cloud & DevOps Knowledge: Experience working with Kubernetes, Docker, Terraform, CI/CD pipelines, and observability tools.
- Database Monitoring: Knowledge of SQL and NoSQL databases and monitoring their performance.
- Strong analytical and problem-solving skills for detecting and resolving performance issues.
- Excellent communication and collaboration skills to work across teams (IT Ops, DevOps, Security, Application teams).