Search by job, company or skills

Manager, Data Operations & Management - Vertex AI

  • Posted 19 minutes ago
  • Be among the first 10 applicants

Job Description

About McDonald's:

One of the world's largest employers with locations in more than 100 countries, McDonald's Corporation has corporate opportunities in Hyderabad. Our global offices serve as dynamic innovation and operations hubs, designed to expand McDonald's global talent base and in-house expertise. Our new office in Hyderabad will bring together knowledge across business, technology, analytics, and AI, accelerating our ability to deliver impactful solutions for the business and our customers across the globe.

Work location: Hyderabad, India

Work pattern: Full time role.

Work mode: Hybrid.

Position Summary:

Platform Engineer III:

We are seeking an experienced AI/ML Platform Engineer to lead the design, engineering, automation, and governance of enterprise-scale AI/ML platforms on Google Cloud Platform (GCP). The ideal candidate will be responsible for building and evolving secure, scalable, and self-service AI/ML platform capabilities that enable Data Scientists, ML Engineers, and Product Teams to accelerate the development and deployment of AI solutions.

This role requires strong expertise in cloud platform engineering, Infrastructure as Code (IaC), platform reliability, AI/ML ecosystem services, security, governance, and developer experience. The candidate will play a key role in driving platform strategy, architecture standards, automation, and operational excellence across the AI/ML landscape.

Who we're looking for:

Primary Responsibilities:

AI/ML Platform Architecture & Engineering:

  • Design, build, and maintain enterprise-grade AI/ML platform services leveraging Vertex AI and GCP.
  • Define platform architecture standards, reusable patterns, and reference implementations for AI/ML workloads.
  • Develop scalable self-service capabilities for model development, training, deployment, and lifecycle management.
  • Establish platform engineering best practices to improve scalability, availability, security, and operational efficiency.
  • Lead platform modernization initiatives and adoption of emerging AI technologies including GenAI and LLM platforms.

Platform Automation & Infrastructure Engineering:

  • Build and manage Infrastructure as Code (IaC) solutions using Terraform and cloud-native automation frameworks.
  • Automate provisioning of AI/ML environments, networking, security policies, Vertex AI resources, and cloud services.
  • Develop reusable automation modules, deployment blueprints, and platform accelerators.
  • Drive standardization of cloud resources and platform deployments across multiple environments and business units.

AI/ML Platform Reliability & Operations:

  • Define and implement platform reliability, resiliency, disaster recovery, and high-availability strategies.
  • Establish SLAs, SLOs, and operational metrics for platform services.
  • Build proactive monitoring, observability, alerting, and incident management capabilities.
  • Lead root cause analysis and platform optimization initiatives for critical incidents and service disruptions.
  • Drive continuous improvement programs focused on platform performance and operational excellence.

CI/CD and Platform Engineering Excellence:

  • Design and govern enterprise CI/CD frameworks supporting AI/ML platform services.
  • Implement automated testing, validation, security scanning, compliance checks, and release automation.
  • Integrate GitHub, GitHub Actions, Cloud Build, Jenkins, SonarQube, and DevSecOps controls into platform pipelines.
  • Promote engineering best practices including code quality, versioning, release governance, and platform standards.

Security, Governance & Compliance:

  • Implement enterprise security controls, governance frameworks, and policy-driven platform automation.
  • Design and manage access controls, IAM policies, service identities, and secrets management.
  • Ensure compliance with organizational standards, regulatory requirements, and data governance policies.
  • Partner with cybersecurity, architecture, and risk teams to establish secure AI/ML platform foundations.
  • Enable Responsible AI, auditability, model governance, and compliance reporting capabilities.

AI Platform Services & Ecosystem Enablement:

  • Lead implementation and governance of Vertex AI services including Workbench, Pipelines, Feature Stores, Model Registry, Endpoints, and GenAI capabilities.
  • Build and optimize integrations across BigQuery, Cloud Storage, Dataflow, Dataproc, Cloud Run, Cloud Functions, Pub/Sub, and Dataplex.
  • Drive platform enhancements supporting MLOps, LLMOps, and AI application development.
  • Enable platform capabilities that simplify onboarding and accelerate AI solution delivery.

Leadership & Stakeholder Management:

  • Provide technical leadership and mentorship to engineers and platform teams.
  • Collaborate with Data Scientists, ML Engineers, Architects, Product Owners, and Business Stakeholders.
  • Lead architecture reviews, technical design discussions, and platform roadmap planning.
  • Drive adoption of platform engineering principles and best practices across the organization.
  • Influence strategic decisions related to AI/ML platform capabilities and cloud modernization.

Skills:

  • 7-11 years of experience in Platform Engineering, Cloud Engineering, DevOps, Site Reliability Engineering, or AI/ML Platform Engineering.
  • 4+ years of hands-on experience designing and implementing AI/ML platforms on Google Cloud Platform.
  • Strong experience with Vertex AI and enterprise AI/ML ecosystem services.
  • Expertise in Infrastructure as Code using Terraform.
  • Strong programming experience in Python, SQL, and automation frameworks.
  • Hands-on experience with GitHub Enterprise, GitHub Actions, Jenkins, Cloud Build, SonarQube, and CI/CD implementations.
  • Experience with BigQuery, Cloud Storage, Dataflow, Dataproc, Cloud Run, Cloud Functions, Pub/Sub, and Dataplex.
  • Strong understanding of cloud networking, security architecture, IAM, governance, and compliance controls.
  • Experience implementing observability solutions using Cloud Monitoring, Logging, dashboards, and telemetry frameworks.
  • Proven experience designing scalable, reliable, and secure cloud platforms.
  • Excellent problem-solving, communication, stakeholder management, and leadership skills

Additional Information:

McDonald's is committed to providing qualified individuals with disabilities with reasonable accommodations to perform the essential functions of their jobs. McDonald's provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to sex, sex stereotyping, pregnancy (including pregnancy, childbirth, and medical conditions related to pregnancy, childbirth, or breastfeeding), race, color, religion, ancestry or national origin, age, disability status, medical condition, marital status, sexual orientation, gender, gender identity, gender expression, transgender status, protected military or veteran status, citizenship status, genetic information, or any other characteristic protected by federal, state or local laws. This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation and training.

McDonald's Capability Center India Private Limited (McDonald's in India) is a proud equal opportunity employer and is committed to hiring a diverse workforce and sustaining an inclusive culture. At McDonald's in India, employment decisions are based on merit, job requirements, and business needs, and all qualified candidates are considered for employment. McDonald's in India does not discriminate based on race, religion, colour, age, gender, marital status, nationality, ethnic origin, sexual orientation, political affiliation, veteran status, disability status, medical history, parental status, genetic information, or any other basis protected under state or local laws.

Nothing in this job posting or description should be construed as an offer or guarantee of employment.

More Info

Job Type:
Industry:
Employment Type:

Job ID: 153614241

Similar Jobs

Hyderabad, India

Skills:

PythonValuation controlsquantitative modelingRisk management

Hyderabad, India

Skills:

GithubAutomated TestingCursorDockerGitlabAzure DevOpsrelease managementJenkinsGitDevSecOpsBitbucketKubernetesInfrastructure as CodeMetadata APIGitOpsdeployment automationClaudesecurity integrationGitHub ActionsSalesforce CLICircleCIobservabilitySalesforce DXCopadoAzure ReposSalesforce DevOpsGitHub CopilotGitLab CIGeminiPlatform Engineering

Hyderabad, India

Skills:

react.js SqlNode.jsRDBMSDockerGitTypescriptJavascriptCI CD

Hyderabad, India

Skills:

distributed caching JavaRDBMSPythonNoSQL DBs

Hyderabad, India

Skills:

.Net 10Power BiPower AutomateSqlAngularReactPythonPower AppsLLM modelAzure servicesMicrosoft Power PlatformNoSQL Database

Beware of Scammers

We don’t charge money for job offers