Search by job, company or skills

Cloud & AI Infrastructure Engineer (AWS)

Early Applicant
  • Posted a day ago
  • Be among the first 10 applicants

Job Description

We are hiring for one of our Client/Startup

The Role

You will own the cloud infrastructure that powers a production SaaS platform serving financial

institutions. This is not a passive keep-the-lights-on role — you will architect scaling strategies,

optimise costs, harden security, and build the infrastructure backbone for AI workloads. As an

early hire at a high-growth startup, you will also contribute to application development across

the stack when needed.

Primary Responsibilities

- Cloud Infrastructure Ownership — Manage and evolve the AWS production

environment: compute, networking, storage, databases, caching, and security services

-Scaling & Performance — Design and implement auto-scaling strategies, diagnose

production bottlenecks, and ensure the platform handles growing client workloads

reliably

-Infrastructure as Code — Own and extend the Terraform codebase for repeatable,

auditable infrastructure deployments

-AI/ML Infrastructure — Support and scale AI workloads including LLM service

deployments, model inference pipelines, and vector storage

-CI/CD & Deployment — Maintain and improve automated build, test, and deployment

pipelines across backend and frontend services

-Observability & Incident Response — Build and maintain monitoring, alerting, and

distributed tracing; lead incident diagnosis and resolution

-Security & Compliance — Uphold security best practices for a FinTech platform,

including encryption, access controls, WAF management, and compliance readiness

(SOC2, AWS FTR)

-Cost Optimisation — Monitor cloud spend, right-size resources, and implement

cost-saving strategies as the platform scales.

Secondary Responsibilities

-Contribute to backend and/or frontend application development as needed - -

-Collaborate with business and operations teams on client onboarding infrastructure

-Evaluate and integrate new AWS services or third-party tools that benefit the platform.

What We're Looking For

Must Have

-3+ years hands-on experience with AWS in a production environment

-Strong experience with containerised workloads (ECS/Fargate or EKS, Docker, ECR)

-Proficiency in Infrastructure as Code (Terraform strongly preferred)

-Experience with CI/CD pipelines (GitHub Actions, GitLab CI, or similar)

-Solid understanding of networking (VPC, subnets, security groups, load balancers,

DNS)

-Experience with managed databases (RDS PostgreSQL, ElastiCache/Redis)

-Familiarity with monitoring and observability tools (CloudWatch, X-Ray, Grafana, or

similar)

-Working knowledge of Linux, scripting (Bash/Python), and troubleshooting production

issues

-Understanding of security fundamentals: IAM, encryption at rest/in transit, secrets

management

Great to Have

-Experience scaling AI/ML workloads in production (model serving, GPU instances,

managed AI services like Bedrock/SageMaker)

-Backend development experience with Python (Django/FastAPI)

-Frontend development experience (React/Next.js)

-Experience with asynchronous task processing (Celery, SQS, or similar)

- Exposure to FinTech or regulated environments (SOC2, FTR, compliance

frameworks)

-Experience with S3 lifecycle management, CloudTrail, GuardDuty, or Security Hub

-Familiarity with serverless patterns (Lambda, Amplify)

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 151835451

Beware of Scammers

We don’t charge money for job offers