Search by job, company or skills

5-7 Years
Early Applicant
  • Posted 4 hours ago
  • Be among the first 10 applicants

Job Description

  • Design, develop, and maintain efficient and scalable data pipelines using Python, PySpark, and Databricks on AWS.
  • Work closely with data scientists, analysts, and other engineers to ensure smooth data integration and high-quality data processing.
  • Build and optimize complex ETL (Extract, Transform, Load) workflows to handle large datasets.
  • Implement data ingestion and transformation logic using PySpark on Databricks to improve data processing performance.
  • Ensure data quality, accuracy, and consistency across all data pipelines.
  • Collaborate with cross-functional teams to identify and resolve bottlenecks in data systems.
  • Utilize AWS cloud services (e.g., S3, Redshift, RDS, EMR) for data storage and management.
  • Write SQL queries to extract, manipulate, and transform data stored in relational databases.
  • Develop, deploy, and maintain APIs for integrating data systems and enabling data access.
  • Monitor and troubleshoot data pipelines, ensuring they operate smoothly with minimal downtime.
  • Stay up to date with the latest industry tools, technologies, and best practices in cloud data engineering and big data processing.

Job Category: Data Services

Job Type: Full Time

Job Location: Pune Gurgaon noida

Experience: 5+ Years

  • Design, develop, and maintain efficient and scalable data pipelines using Python, PySpark, and Databricks on AWS.
  • Work closely with data scientists, analysts, and other engineers to ensure smooth data integration and high-quality data processing.
  • Build and optimize complex ETL (Extract, Transform, Load) workflows to handle large datasets.
  • Implement data ingestion and transformation logic using PySpark on Databricks to improve data processing performance.
  • Ensure data quality, accuracy, and consistency across all data pipelines.
  • Collaborate with cross-functional teams to identify and resolve bottlenecks in data systems.
  • Utilize AWS cloud services (e.g., S3, Redshift, RDS, EMR) for data storage and management.
  • Write SQL queries to extract, manipulate, and transform data stored in relational databases.
  • Develop, deploy, and maintain APIs for integrating data systems and enabling data access.
  • Monitor and troubleshoot data pipelines, ensuring they operate smoothly with minimal downtime.
  • Stay up to date with the latest industry tools, technologies, and best practices in cloud data engineering and big data processing.

More Info

Job Type:
Industry:
Function:
Employment Type:

Job ID: 153566829

Similar Jobs

Pune, India

Skills:

snowflake S3RedshiftJenkinsLambdaSpark StreamingApache KafkaGitlabData GovernancePythonAWSAlationGlueData Quality Mesh

Pune, India

Skills:

.NETdata warehouses Spark SQLPower BiAzure Log AnalyticsPowershell ScriptingHiveAzure Data FactoryData lakesMicrosoft Azure Data platformAzure SQL Data WarehouseAzure Storage ServicesAzure Application InsightsData BricksStream Analyticsdata martsEvent HubsAzure SQL DBAzure Analysis Services

Pune, India

Skills:

JavaUnixApache FlinkData ModelingSchema DesignPysparkData CleansingApache SparkData Warehouse ConceptsShell ScriptingSqlELTApache AirflowLinuxApache KafkaRestful ApisPythonEtlData TransformationData Quality Validation

Pune, India

Skills:

snowflake KafkaELTDevopsSparkIncident ManagementDatabricksRestful ApisAzureEtlAWSAirflowMonitoring

Pune, India

Skills:

proxmox ServicenowKvmWindows ServerPrometheusGrafanaDatadogTerraformVcenterPythonAWSEsxiVmware VspherePowerShellBashJiraUbuntuGitNsxAnsibleRhelPowercliVsannetworking fundamentals

Beware of Scammers

We don’t charge money for job offers