

Search by job, company or skills
Role Overview
We are looking for a Lead Data Engineer with strong experience building scalable batch, near-real-time, and streaming data platforms on Microsoft Azure.
The role requires hands-on expertise in Python, PySpark, advanced SQL, Azure data services, Medallion Architecture, and deploying Apache Spark workloads on Kubernetes or Azure Kubernetes Service (AKS).
.Key Responsibilities
Required Skills
Job ID: 153637963
Skills:
Api Management, Apache Spark, Azure Databricks, Azure Data Factory, Terraform, Delta Lake Architecture, Azure Data Lake Storage Gen2, Service Principals, Azure Key Vault, Unity Catalog, Databricks Asset Bundles, rbac, CI CD pipelines, Event Grid, Azure Data Services
Skills:
Csv, Pyspark, Json, ELT, Git, Docker, MongoDB, Rest Apis, Advanced Sql, Kubernetes, Python, Etl, OLAP engines
Skills:
Data Lineage, React, Typescript, Javascript, Python, Pipeline Builder, AIP toolset, Palantir Foundry, OSDK
Skills:
Apache Spark, Kafka, Apache Airflow, Pandas, Elasticsearch, MongoDB, pgvector, Polars, Vector databases, Qdrant, NoSQL databases, Data models and storage architectures, Distributed data systems, Data processing performance and optimisation, Data pipelines, Milvus
Skills:
Spark SQL, S3, RDS, Aws Services, Pyspark, Emr, Jenkins, Git, Devops Tools, Python, Big Data concepts, AWS architecture, AWS cloud microservices, cloud implementation projects, Aurora, Snowflake SQL, Glue, Data Vault 2.0