

Search by job, company or skills

Note: - This is 6 Months Contractual role based on projects chances of extension is 90% for 12 months further with 100% remote opportunity.
We are looking for a highly experienced Azure Databricks Data Engineer with 8+ years of overall experience and strong hands-on expertise in Azure Databricks, PySpark, Databricks Asset Bundles (DAB), Azure DevOps, and CI/CD.
The ideal candidate will be responsible for designing and developing scalable data engineering solutions, implementing robust ETL/ELT pipelines, and automating Databricks deployments across multiple environments using DevOps best practices.
Key Responsibilities
Required Skills
Must Have
Job ID: 151812945
Skills:
Devops, Python, GitHub Copilot, Spec Kit development, AI models
Skills:
Hadoop, Data Modeling, Pyspark, Apache Spark, Kafka, Node.js, Numpy, Pandas, Docker, Flask, FastAPI, Azure, Kubernetes, AWS, LangChain, GenAI LLMs, Databricks Delta Lake, Python Ecosystem, Pinecone, Milvus, LlamaIndex
Skills:
graph databases , snowflake , Git, Python, Sql, Azure DevOps, AWS cloud services, vector databases
Skills:
Pyspark, Data Modeling, ELT, Cloud Storage, Numpy, Docker, Terraform, Python, BigQuery, Sql, Jenkins, Pandas, Spark, Kubernetes, Etl, GCP Core Services, Infrastructure as Code, Cloud Dataflow, Real-time streaming architectures, CI CD, Pub Sub, Cloud Run, GitHub Actions, API integrations, Cloud Functions, Cloud Dataproc, Cloud Composer Airflow
Skills:
Advanced Sql, Python, Analytics Hub data sharing, Dataflow batch streaming pipelines, Batch and streaming data processing, BigQuery advanced optimization and tuning, Scripting and automation, Dataform SQL-based transformations, Dataplex Data Catalog, Data lakehouse architecture concepts, Google Cloud Platform GCP, IAM Security controls, Pub Sub event-driven architectures, Cloud Monitoring Logging, Cloud Storage GCS, Data ingestion frameworks config-driven approach preferred, ETL ELT pipeline design and implementation, Data modeling and warehousing principles, Java or Scala preferred, CDC and incremental processing