Search by job, company or skills

Senior Site Reliability Engineer

Senior Site Reliability Engineer

Akamai Technologies
5-7 Years
Not Disclosed
  • Posted 9 hours ago
  • Be among the first 10 applicants

Job Description

Job Description

Are you passionate about cutting edge technology

Would you like an opportunity to effect change at a leading technology organization

Join our Site Reliability team

The Akamai Cloud Technology Engineering team owns, develops and manages the solutions used by our engineers globally. These solutions help to run one of the largest distributed systems in the world. We work closely with internal teams and stakeholders to create innovative, powerful, scalable, highly reliable and secure systems at scale.

Partner with the best

This Senior Site Reliability Engineer (Linux) role involves improving automation and efficiency for internal teams while ensuring operational excellence. Responsibilities include enhancing system reliability, scalability, and performance by designing and maintaining infrastructure, tools, and processes. Collaborate to address technical challenges, optimize deployments, and support applications. Drive continuous improvement through workflow automation, system monitoring, and performance optimization to achieve organizational objectives.

As a Senior Site Reliability Engineer, you will be responsible for:

  • Developing processes, plans, and infrastructure to deploy new software components and updates safely and efficiently at scale
  • Improving our system monitoring and analysis platform to speed error detection and remediation, enhancing performance and reliability
  • Improving our system monitoring and analysis platform to speed error detection and remediation, enhancing performance and reliability.

Do What You Love

To be successful in this role you will:

  • Have 5+ years of relevant experience and a Bachelors degree in Computer Science or related field
  • Specialize in Linux administration, demonstrating expertise in Python and Bash scripting languages for advanced systems management and automation tasks.
  • Utilize Salt Stack, Ansible, Terraform for infrastructure automation, alongside CI/CD tools including Jenkins for streamlined deployment and management.
  • Demonstrate expertise with observability or monitoring tools like Prometheus, Grafana, ELK/Open Search, Datadog, and Splunk, Docker
  • Have hand-on mastery with Kubernetes experience with any cloud platform, such as AWS, GCP, Azure, or an equivalent alternative

About Us

At Akamai, we make life better for billions of people, trillions of times a day.

Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.

Our Focus Is Simple

Cloud and Edge: Running apps closer to users for instant performance.

Security: Neutralizing threats before they ever reach your data.

Content Delivery: Scaling the world's biggest moments without a glitch.

AI: Enabling our customers to build, secure, and scale AI apps on the world's most distributed cloud platform.

At Akamai, we don't just support the internet; we power and protect it, because behind every great digital experience is a massive hidden challenge. And we're the ones who solve it. When millions of people hit play or pay, Akamai ensures it just works.

Benefits at Akamai: We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job's needs

Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.

We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

Connect with us on social and see what life at Akamai is like!

More Info

Job Type:
Industry:
Function:
Employment Type:

About Company

Similar Jobs

6-8 yrs
Bengaluru, India
Skills:
C#, C++, Distributed Systems, PowerShell, Networking, Python, tools automation scripting, troubleshooting debugging, KQL, telemetry-based analysis
3-5 yrs
Bengaluru, India
Skills:
Unix, C, Continuous Integration, Software Architecture, Javascript, Linux, Distributed Systems, Python, Disaster Recovery Planning, alerting tools, Root Cause Analysis, Compliance, performance optimization techniques, scalability design patterns, Troubleshooting, continuous delivery automation, security standards
5-7 yrs
Bengaluru, India
Skills:
Elk, PostgreSQL, Prometheus, Grafana, Terraform, MySQL, Python, AWS, RDS, Redis, Gcp, Ansible, Load Balancing, HashiCorp Vault, Go, OpenTelemetry, Cloud SQL, Linux systems, Loki, Kubernetes EKS, Google Cloud Operations, Google Secret Manager, GitLab CI, Memorystore, Spinnaker, CI CD pipelines, AWS Secrets Manager, ArgoCD, DNS routing
8-12 yrs
Bengaluru, India
Skills:
Saml, Prometheus, Kafka, Spring Boot, Datadog, Docker, Terraform, Teamcity, Python, AWS, Oauth, Java, RDS, Sso, Jenkins, Cloudwatch, Bitbucket, Sqs, Helm, Kubernetes, Go, Aurora, GitHub Actions, OpenTelemetry, FluxCD
8-12 yrs
Bengaluru, India
Skills:
Incident Management, Infrastructure optimization, Forecasting, Cloud cost optimization, AWS billing analysis, FinOps, Budget Tracking, Production Support, Capacity Planning, Linux systems knowledge, Performance Analysis, Tagging, cost allocation, Reliability improvements