
Search by job, company or skills
Our client is a high-scale global consumer platform serving millions of users daily. They are rebuilding their data foundation as an AI-native platform to power the next generation of programmatic advertising, personalized discovery, and marketplace intelligence.
We are hiring a hands-on Staff / Lead Data Engineer (AI-Native) to architect petabyte-scale systems that feed real-time bidding, recommendation models, and LLM-powered analytics products.
Role Mandate: 80% Hands-on Engineering / 20% Technical Leadership
This is NOT a pure management role. You will be a player-coach - designing and shipping critical pipelines yourself while mentoring a small pod of engineers. Strong individual contributors with no formal people management experience but with deep technical depth are strongly encouraged to apply.
What You Will Build
1. AI-Native Data Foundation at Scale
Design, build, and operate AI-ready batch + streaming ETL/ELT pipelines ingesting 100TB+ daily from ad servers, mobile SDKs, transactional systems, and 3rd-party APIs. Build for LLM and ML consumption from day one.
2. Real-Time & Agentic Data Systems
Develop low-latency streaming jobs using Spark Streaming, Flink, or Kafka Streams for real-time use cases: fraud detection, bid optimization, dynamic pricing, and real-time personalization. Enable online inference and agentic decisioning.
3. Lakehouse for AI & Analytics
Model and optimize massive datasets on a modern lakehouse [Databricks / Snowflake / BigQuery] to serve BI, embedded analytics, and AI/ML workloads with a focus on performance, cost, and feature freshness.
4. Data Products for AI
Build reliable data products for Data Science & ML: feature stores [Feast / Tecton], vector stores for semantic search & RAG, training datasets with point-in-time correctness, and online-offline parity.
5. Reliability for Tier-0 AI Systems
Own data quality, observability, anomaly detection, and lineage for Tier-0 datasets that directly power revenue, user experience, and model performance.
Required Qualifications
Preferred - You Stand Out If You Have:
Job ID: 152246787
Skills:
snowflake , Spark, Kafka, Databricks, Kubernetes, Python, Sql, Airflow, dbt
Skills:
bedrock , Apache Flink, Gcp, Docker, Openshift, Azure, Kubernetes, Python 3, AWS, CrewAI, Ray, Google BigQuery, Langgraph, LiteLLM, VertexAI