Search by job, company or skills

Senior Data Engineer

8-10 Years
  • Posted 16 hours ago
  • Be among the first 10 applicants

Job Description

Note: - This is 6 Months Contractual role based on projects chances of extension is 90% for 12 months further with 100% remote opportunity.

We are looking for a highly experienced Azure Databricks Data Engineer with 8+ years of overall experience and strong hands-on expertise in Azure Databricks, PySpark, Databricks Asset Bundles (DAB), Azure DevOps, and CI/CD.

The ideal candidate will be responsible for designing and developing scalable data engineering solutions, implementing robust ETL/ELT pipelines, and automating Databricks deployments across multiple environments using DevOps best practices.

Key Responsibilities

  • Design, develop, and maintain enterprise-scale data pipelines using Azure Databricks and PySpark.
  • Develop scalable ETL/ELT solutions using PySpark, Spark SQL, and SQL.
  • Build and optimize Delta Lake tables and Lakehouse data processing workflows.
  • Implement and manage Databricks Asset Bundles (DAB) for packaging, configuration, and deployment of Databricks resources.
  • Design and maintain Azure DevOps CI/CD pipelines for Databricks deployments.
  • Automate deployment of Databricks notebooks, jobs, workflows, and configurations across Dev, QA, UAT, and Production environments.
  • Integrate Databricks development with Git, GitHub, or Azure Repos.
  • Implement environment-specific configurations and deployment strategies.
  • Work with Azure Data Lake Storage Gen2, Azure Key Vault, Azure Data Factory, and other Azure services.
  • Develop reusable PySpark frameworks and production-ready data processing solutions.
  • Perform Spark performance tuning, including optimization of partitions, joins, caching, file sizes, and Delta Lake operations.
  • Implement data quality, validation, logging, monitoring, and error-handling frameworks.
  • Troubleshoot complex production issues and provide technical solutions.
  • Collaborate with architects, data engineers, DevOps teams, and business stakeholders.
  • Provide technical guidance and mentoring to junior and mid-level engineers.

Required Skills

Must Have

  • 8+ years of overall experience in Data Engineering / Big Data / Cloud Data Engineering.
  • Strong hands-on experience with Azure Databricks.
  • Strong expertise in PySpark and Apache Spark.
  • Hands-on experience with Databricks Asset Bundles (DAB).
  • Strong experience with Azure DevOps and CI/CD pipelines.
  • Experience automating Databricks deployments across multiple environments.
  • Strong knowledge of Delta Lake and Lakehouse architecture.
  • Strong SQL skills.
  • Experience with Git / GitHub / Azure Repos.
  • Experience with Azure Data Lake Storage Gen2 (ADLS Gen2).
  • Strong understanding of ETL/ELT concepts and modern data engineering practices.

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 151812945

Similar Jobs

Chennai, India

Skills:

DevopsPythonGitHub CopilotSpec Kit developmentAI models

Bengaluru, India

Skills:

HadoopData ModelingPysparkApache SparkKafkaNode.jsNumpyPandasDockerFlaskFastAPIAzureKubernetesAWSLangChainGenAI LLMsDatabricks Delta LakePython EcosystemPineconeMilvusLlamaIndex

Chennai

Skills:

graph databases snowflake GitPythonSqlAzure DevOpsAWS cloud servicesvector databases

Hyderabad, India

Skills:

PysparkData ModelingELTCloud StorageNumpyDockerTerraformPythonBigQuerySqlJenkinsPandasSparkKubernetesEtlGCP Core ServicesInfrastructure as CodeCloud DataflowReal-time streaming architecturesCI CDPub SubCloud RunGitHub ActionsAPI integrationsCloud FunctionsCloud DataprocCloud Composer Airflow

Chennai, India

Skills:

Advanced SqlPythonAnalytics Hub data sharingDataflow batch streaming pipelinesBatch and streaming data processingBigQuery advanced optimization and tuningScripting and automationDataform SQL-based transformationsDataplex Data CatalogData lakehouse architecture conceptsGoogle Cloud Platform GCPIAM Security controlsPub Sub event-driven architecturesCloud Monitoring LoggingCloud Storage GCSData ingestion frameworks config-driven approach preferredETL ELT pipeline design and implementationData modeling and warehousing principlesJava or Scala preferredCDC and incremental processing

Beware of Scammers

We don’t charge money for job offers