

Search by job, company or skills

Data Pipeline Development : Design, develop, and optimize data pipelines to ingest, process, and transform data from various sources (e.g., APIs, databases, into the data warehouse.
Data Integration: Integrate data from various structured and unstructured sources into the Databricks Lakehouse environment, ensuring data accuracy and reliability
Data Lakehouse storage Management: Design and maintain data warehouse solutions using medallion architecture practices, optimizing storage, cloud utilization, costs and query performance
Collaboration with Data Teams : Work closely with data scientists, analysts, to understand requirements, translate them into technical solutions, and implement data solutions.
Data Quality and Monitoring : Cleanse, transform, and enrich data. Implement data quality checks and establish monitoring processes to ensure data integrity and accuracy. Implement monitoring for data pipelines and troubleshoot any issues or failures promptly to ensure data reliability.
Optimization and Performance Tuning: Optimize data processing workflows for performance, reliability, and scalability, including tuning spark jobs, caching, and partitioning data appropriately.
Data Security and Privacy: Manage and organize data lakes using Unity catalog, ensuring proper governance, security, role-based access and compliance with data management policies
Preferred Qualifications:
Technical Skills:
Godrej Agrovet is a diversified, Research Development focused agri-business company, dedicated to improving the productivity of Indian farmers by innovating products and services that sustainably increase crop and livestock yields. We hold leading market positions in the different businesses in - Animal Feed, Crop Protection, Oil Palm, Dairy and Poultry and Processed Foods.
Job ID: 113100921
Skills:
Power Bi, Scala, Azure Databricks, Data Warehouse, Sql, Azure Data Factory, Spark, Python, Azure DevOps, Azure Data Lake Storage, Synapse Analytics, Delta Lake
Skills:
snowflake , operational support , Metadata Management, Sql, ELT, Data Modeling, Data Governance, Etl, Data Quality, Python, RAG pipelines, vector databases, AI LLM-enabled data architectures
Skills:
Java, Cassandra, Scala, Big Data, Kafka, Sql, Nosql, Hive, RDBMS, Presto, Spark, MongoDB, Python, HDFS
Skills:
snowflake , Sql, ELT, Etl, Data Modeling, Data Governance, Python, Metadata Management, data quality controls
Skills:
Sql, Azure Data Factory, Power Bi, Azure Databricks, Python, Bi Tools, Scala, Spark, Azure DevOps, Delta Lake, Azure Data Lake Storage, Synapse Analytics, CI CD practices