
Search by job, company or skills
We are seeking a highly skilled Site Reliability Engineer (SRE) to maintain the stability of our software product throughout its entire development lifecycle. You will be responsible for measuring and monitoring the system's general state, analyzing incident data, automating monitoring processes, and developing frameworks and scripts to enhance product stability and reliability.
Roles & Responsibilities:
Skills Required:
QUALIFICATION:
Our services deliver a total solution package designed to meet our clients' complete business and technology needs.
With recognized technology and industry expertise, our dedicated professionals focus on understanding clients' competitive challenges and implementing the right solutions to achieve their strategic objectives. At APEX, we strive to build long-term, committed partnerships with our clients, helping them every step of the way on their path to success. APEX approaches every engagement with one objective in mind - to help our client win and grow.
Job ID: 122964613
Skills:
Servicenow, Grafana, Terraform, Ansible, Dynatrace, Microsoft Azure, Splunk, Itil Processes, Kubernetes, Python, GKE, DevOps practices, AKS, GitHub Actions, microservices architecture, ITSM tools, API-based architectures
Skills:
Elk, Prometheus, Bash, Grafana, Datadog, Jenkins, Gcp, Linux, Docker, Terraform, Ansible, Azure, Kubernetes, Python, AWS, Go, CI CD, GitLab CI, GitHub Actions, OpenTelemetry
Skills:
Apache Flink, Prometheus, Grafana, Python, Kubernetes, Confluent Kafka, AlertManager, GitHub Actions, AWS-MSK, Azure Event Hub, Confluent Cloud
Skills:
New Relic, Azure Functions, Dynatrace, Microsoft Azure, Datadog, Azure Service Bus, Azure Alerts, Azure Portal, Azure Monitor, Application Insights, Log Analytics
Skills:
Google Cloud Platform, Linux Administration, Automation, Kubernetes, Incident Management, Golang, Docker, Dynatrace, Production Support