Search by job, company or skills

PySpark Developer

PySpark Developer

Infosys Limited
Fresher
Not Disclosed
  • Posted 6 hours ago
  • Be among the first 10 applicants

Job Description

Responsibilities :

Design, develop, and maintain data pipelines using PySpark. Develop ETL/ELT processes for ingesting, transforming, and loading large volumes of data. Write optimized PySpark code for data processing and transformation. Work with structured and semi-structured data from multiple sources. Develop and optimize SQL queries for data extraction and validation. Troubleshoot data quality and performance issues. Collaborate with Data Engineers, Analysts, and Business teams to understand requirements. Participate in code reviews and follow data engineering best practices. Monitor and support production data pipelines.

Additional Responsibilities:

Exposure to cloud platforms such as AWS, Azure, or GCP. Knowledge of Databricks. Experience with workflow orchestration tools such as Airflow. Understanding of CI/CD concepts.

Technical and Professional Requirements:

Strong experience in PySpark. Good programming knowledge of Python. Hands-on experience with SQL. Understanding of ETL/ELT concepts and data warehousing. Experience working with large datasets and distributed processing. Knowledge of Spark SQL, DataFrames, and Spark transformations. Familiarity with Linux/Unix environment.

More Info

Job Type:
Industry:
Employment Type:

Key Skills

Big Data - Data Processing

About Company

Similar Jobs

Bengaluru, India
Skills:
Big Data - Data Processing, Pyspark, Technology