

Search by job, company or skills

This job is no longer accepting applications
Role & responsibilities
We are seeking a highly skilled Data Engineer with deep expertise in PySpark and the Cloudera Data Platform (CDP) to join our data engineering team. As a Data Engineer, you will be responsible for designing, developing, and maintaining scalable data pipelines that ensure high data quality and availability across the organization. This role requires a strong background in big data ecosystems, cloud-native tools, and advanced data processing techniques.
The ideal candidate has hands-on experience with data ingestion, transformation, and optimization on the Cloudera Data Platform, along with a proven track record of implementing data engineering best practices. You will work closely with other data engineers to build solutions that drive impactful business insights.
Responsibilities
Preferred candidate profile
Technical Skills
Soft Skills
Clover Infotech is a leading global IT services and consulting company. We provide solutions and services across application and technology modernization, cloud enablement, data management, automation, and assurance services. Clover Infotech is among the most preferred Oracle Partners with extensive experience in implementation and management of Oracle Fusion Applications and Oracle Cloud Infrastructure (OCI).
Job ID: 108697653
Skills:
Hadoop, Pyspark, Spark, Data Governance, Sql, Python, ELT, Etl, Security, compliance frameworks
Skills:
data engineering , stream processing , Pyspark, Databases, MYSQL, Tableau, JIRA, Sql, Unit Test, Git, End To End Testing, Confluence, Debugging, Databricks, AI tools, data pipelines, Databricks dashboarding, Optimization Skills, key-value stores, blob stores, delta live tables
Skills:
HBase, Sql, Etl, Hive, Impala, Apache Oozie, Hadoop, Pyspark, Kafka, Airflow, Cloudera Data Platform, HDFS
Skills:
Hadoop, Pyspark, Apache Spark, Sql, Mapreduce, Apache Airflow, Nosql, Jenkins, Git, Hive, Pandas, Oozie, Oracle, Python, Etl
Skills:
Hive, Hadoop, Pandas, Pyspark, Python, Sql, Mapreduce, NoSQL DBMS, Jupyter