Search by job, company or skills

10-20 Years
Not Disclosed
Quick Apply
  • Posted a month ago
  • Over 300 applicants have applied

Job Description

We are looking for a highly experienced Solution Architect – Databricks to work closely with enterprise customers in designing, developing, optimizing, and supporting scalable data engineering and analytics solutions on the Databricks platform.

The ideal candidate should have strong hands-on expertise in Databricks, Apache Spark, PySpark, distributed computing, cloud platforms, performance optimization, and solution architecture. This is a highly technical and client-facing role requiring the ability to independently lead architecture discussions, troubleshoot complex Databricks/Spark issues, and provide implementation guidance.

Key Responsibilities

  • Design and implement scalable Databricks Lakehouse solutions.
  • Define end-to-end data engineering and platform architecture.
  • Build and optimize data pipelines using Databricks, Spark, PySpark, SQL, and Delta Lake.
  • Design batch and streaming data-processing solutions.
  • Provide technical guidance on Databricks architecture and best practices.
  • Troubleshoot and optimize complex Spark and Databricks performance issues.
  • Work with Delta Lake, Unity Catalog, Workflows, Auto Loader, Databricks SQL, Lakeflow/DLT, and Serverless.
  • Support enterprise Databricks implementation and modernization initiatives.
  • Design and support CI/CD processes for Databricks deployments.
  • Work with Git, Terraform, Azure DevOps, GitHub, GitLab, or Jenkins.
  • Provide guidance on security, governance, access control, and Unity Catalog.
  • Conduct architecture reviews, code reviews, troubleshooting, and technical mentoring.
  • Work closely with customer architects, engineering teams, and business stakeholders.

Required Skills

  • 10+ years of overall technology/consulting experience.
  • 7+ years of experience in Data Engineering, Big Data, Data Platforms, or Analytics.
  • Strong hands-on experience with Databricks.
  • Experience delivering 6–8+ end-to-end Databricks projects.
  • Strong expertise in Apache Spark and PySpark.
  • Deep understanding of Spark internals including Driver/Executors, DAG, Jobs, Stages, Tasks, Partitioning, Shuffle, Catalyst Optimizer, AQE, and Memory Management.
  • Strong experience in Spark performance tuning, query optimization, data skew, partitioning, and join optimization.
  • Strong knowledge of Delta Lake, Unity Catalog, Databricks Workflows, Auto Loader, and Databricks SQL.
  • Strong ETL/ELT, data pipelines, data modeling, batch and streaming experience.
  • Deep expertise in at least one cloud platform: AWS, Azure, or GCP.
  • Working knowledge of at least one additional cloud platform.
  • Knowledge of Git, CI/CD, Terraform, and Databricks Asset Bundles.
  • Working knowledge of MLflow/MLOps is preferred.
  • Strong customer-facing consulting and communication skills.

Preferred Qualifications

  • Databricks Certified Data Engineer Professional certification.
  • Experience in Databricks migration and modernization.
  • Hadoop-to-Databricks migration experience.
  • Cloud data warehouse-to-Databricks migration experience.
  • Multi-cloud architecture exposure.
  • Unity Catalog implementation experience.
  • Data governance and streaming architecture experience.
  • Strong technical leadership and mentoring experience.

More Info

Job Type:
Function:
Employment Type:

Key Skills

Similar Jobs

10-15 yrs
Bengaluru, India
Skills:
Mqtt, S3, IIOT, Quicksight, Lambda, Redshift, Cybersecurity, digital manufacturing, Manufacturing IT, IoT Core, cloud-native architectures, SCADA, AWS IoT Services, historians, Edge Computing, Industrial Automation, Industry 4.0 frameworks, OT networks, PLCs, OPC-UA, Greengrass, Data Platforms, enterprise systems, TwinMaker, SiteWise, Industrial IoT architecture, Mes
8-10 yrs
Bengaluru, India
Skills:
Collibra, Pyspark, Informatica, Odi, DataStage, Terraform, Spark, Databricks, Sas Di, Talend, Mosaic AI, Databricks SQL, GenAI, Purview, Alation, MLflow, Unity Catalog, Lakeflow Declarative Pipelines, Databricks Asset Bundles, Delta Lake, Serverless Compute, Immuta
12-14 yrs
Bengaluru, India
Skills:
Servicenow, Data Architecture, Microservices, Nlp, Distributed Systems, AWS, Apis, Soap, Devops, REST, Gcp, Mulesoft, Azure, embeddings, GenAI, Field Service Lightning, RAGAS, vector databases, CI CD, observability tools, LangGraph, retrieval systems, Salesforce Service Cloud, LangChain, LLMs, AI evaluation frameworks, event-driven architectures, Ai, RAG
10-12 yrs
Bengaluru, India
Skills:
snowflake , Ml, Itil, MLops, Data Management, Databricks, Microsoft Azure, LLMOps, FinOps, Ai, Agent Ops, Analytics, DataOps
8-10 yrs
Bengaluru, India
Skills:
Kubernetes), Containerization (Docker, Object-oriented programming (OOP), Retrieval-Augmented Generation (RAG), Nlp, Javascript, Flask, Api Management, Python, Sql, FastAPI, Sentiment Analysis, Serverless Computing, Microservices, Text Analytics, Azure AI Search, Generative AI, Vector databases, Gen-AI, Llm, Transformer models