You will work on a live, high-load fleet management platform that connects tens of thousands of vehicles across enterprise fleets worldwide — processing real-time telemetry, powering mobile apps used by drivers and technicians on the ground, and integrating with hardware, firmware, and 20+ data partners. The system runs 24/7, handles genuine scale, and the work is a mix of complex new features, infrastructure modernisation, and keeping production rock-solid. If you want a project where the data is real, the stakes are real, and the engineering problems are interesting — this is it.
Project Overview:
The platform is a multi-tenant SaaS solution for enterprise fleet management — built on a microservices architecture running on AWS across multiple accounts and environments, processing continuous telemetry streams from connected vehicles, and serving web, mobile, and third-party consumers through a set of REST and event-driven APIs. The backend is Python/Django and Golang, the frontend is React, mobile is native iOS (SwiftUI) and Android (Kotlin), and the infrastructure runs on EKS with Terraform-managed IaC, Jenkins/GitHub Actions CI/CD, and a full observability stack (Prometheus, Grafana, Elastic APM, PagerDuty on-call).
Requirements:
10+ years in DevOps / Platform Engineering / SRE roles
Advanced AWS production experience across multiple services and environments