Job Summary
Cloud Infrastructure Engineers design, build, secure, and operate the cloud foundations that enable application and platform teams to run reliably and securely. As a Senior Software Engineer II, you will deliver platform capabilities, guardrails, and shared infrastructure systems primarily in AWS, with additional support for GCP, that meet organizational standards for reliability, security, and governance.
This role spans reliability, security, governance, and enablement. You will contribute to core foundation architecture, including networking, IAM, encryption and key management, and observability, plus organization guardrails and policy enforcement.
You will build IaC platform capabilities by developing features that enforce approved standards with testing, validation, and security controls built in. You will participate in solution design for cloud platform capabilities, translating requirements into implementations that align with reference architectures. You'll collaborate with Platform, SRE, Security, Network, and application and service teams to drive adoption of platform standards and support compliance and audit readiness.
We need engineers who are hands-on, show strong technical judgment, communicate clearly, and bring curiosity and enthusiasm for solving hard infrastructure problems.
Job Duties
- Contribute to and implement cloud infrastructure primarily in AWS, with additional support for GCP, including networking, IAM, encryption and key management, observability, and cloud-native managed services and serverless compute.
- Implement organization guardrails and policies, including org policies and SCPs, account or project provisioning standards, tagging standards, and baseline constraints.
- Develop IaC platform features that enforce approved configurations through standardized Terraform modules, CI/CD workflows, automated testing and validation, and policy enforcement.
- Contribute to technical proposals and architecture decisions for platform initiatives, balancing security, reliability, cost, and maintainability.
- Design and implement cloud solutions that span multiple systems, including secure patterns for cloud managed services and their dependencies.
- Participate in design reviews and high-risk change reviews, keeping implementations aligned with approved patterns.
- Triage incidents and platform issues in Cloud Infrastructure-owned systems and implement durable fixes that prevent repeat failures.
- Support audit readiness for Cloud Infrastructure-owned systems, including gathering evidence, maintaining control inventory, and participating in SOC 2 and PCI activities.
- Support peers through code and design reviews, contributing to quality and shared standards across the team.
Requirements
- Bachelor's Degree or equivalent practical experience.
- 4+ years of experience in cloud infrastructure, platform engineering, SRE, and/or software engineering roles, or equivalent scope and impact.
- Deep experience designing and operating cloud foundations in AWS, with working knowledge of GCP, across several of: networking, IAM, identity federation, encryption and key management, certificate management, secrets management, governance guardrails and policies, observability, DNS, and OU/folder and account/project provisioning.
- Advanced Terraform expertise, including authoring reusable modules, and enforcing standards through automated testing and policy checks.
- Deep experience designing and operating CI/CD pipelines for infrastructure delivery with safe change controls.
- Ability to participate in system architecture discussions and technical design reviews, with a solid understanding of security, reliability, and scalability requirements.
- Strong scripting and automation skills (Python, Bash, or similar), including building internal tooling, automating operational workflows, and integrating with cloud APIs.
- Experience supporting or executing cloud migrations, including migrating production workloads from on-premises environments to AWS.
- Exceptional incident triage and debugging skills for complex, distributed cloud infrastructure problems, including diagnosing issues in high-pressure production environments.
- Collaborative working style with strong written and verbal communication skills.
- Experience operating in regulated environments, including audit-ready defaults and control implementation.
Advanced Requirements
- Professional AWS certifications, or equivalent demonstrated depth (GCP certifications also valued).
- Direct SOC 2 or PCI audit experience, including evidence production and remediation execution.
- Deep AWS expertise, with hands-on experience migrating production workloads from on-premises data centers to AWS.