Search Jobs

Search by job, company or skills

Data Engineer
  • Posted 31 minutes ago
  • Be among the first 10 applicants

Job Description

Project Role : Data Engineer

Project Role Description : Design, develop and maintain data solutions for data generation, collection, and processing. Create data pipelines, ensure data quality, and implement ETL (extract, transform and load) processes to migrate and deploy data across systems.

Must have skills : PySpark

Good to have skills : Amazon Web Services (AWS), DevOps, GitLab CI/CD

Minimum 3 Year(s) Of Experience Is Required

Educational Qualification : 15 years full time education

Summary:

As a Data Engineer, your typical day involves designing, developing, and maintaining comprehensive data solutions that support data generation, collection, and processing activities. You will be responsible for creating efficient data pipelines that facilitate smooth data flow and ensure high data quality. Your role includes implementing extract, transform, and load processes to migrate and deploy data seamlessly across various systems, enabling reliable and scalable data infrastructure to support organizational needs.

Roles & Responsibilities:

  • Expected to perform independently and become an SME.
  • Required active participation/contribution in team discussions.
  • Contribute in providing solutions to work related problems.
  • Collaborate with cross-functional teams to understand data requirements and deliver effective data solutions.
  • Continuously monitor and optimize data pipelines to improve performance and reliability.
  • Document processes and workflows to maintain clear communication and knowledge sharing within the team.
  • Assist junior team members by providing guidance and support to foster their professional growth.

Professional & Technical Skills:

  • Must To Have Skills: Proficiency in PySpark.
  • Good To Have Skills: Experience with Amazon Web Services (AWS), DevOps, GitLab CI/CD.
  • Strong knowledge of data pipeline architecture and best practices for data processing.
  • Experience in building scalable and fault-tolerant ETL workflows.
  • Familiarity with distributed computing frameworks and big data technologies.
  • Ability to troubleshoot and resolve complex data processing issues efficiently.

Additional Information:

  • The candidate should have minimum 3 years of experience in PySpark.
  • This position is based at our MDC5C - SEZ office.
  • A 15 years full time education is required.




More Info

Job Type:
Industry:
Employment Type:

Key Skills

About Company

Similar Jobs

6-10 yrs
Mumbai, India
Skills:
PysparkData ModelingData GovernanceEncryptionAzure SynapseAzure DevOpsSpark SQLData WarehousingAzure SqlMetadata ManagementGitAzure Data FactoryPurviewKey VaultManaged IdentitySecurityPII handlingEvent HubsNotebooksSemantic ModelsADLS Gen2Dataflows Gen2Semantic layer designLineagerbacLakehouseCI CD pipelinesDelta LakeDelta Parquet formatsMicrosoft FabricAzure Data Services
4-6 yrs
Mumbai, India
Skills:
CloudformationPysparkAmazon S3AWS GlueRedshiftSqlDockerTerraformECSSparkPythonAWSEKSAthena
5-7 yrs
Mumbai, India
Skills:
GithubPerformance TuningData FactoryPysparkAzure DevOpsPower BiApache SparkSqlMetadata ManagementData WarehousingDatabricksData PipelinesWarehouseMicrosoft Fabric componentsGovernance practicesAutomated testing and validation frameworksData quality reconciliationDatabricks WorkflowsSemantic ModelsLakehouse ArchitecturesData Integration FrameworksETL ELT SolutionsLakehouseCI CD pipelinesDelta LakeMicrosoft FabricAutomation and Orchestration
7-9 yrs
Mumbai, India
Skills:
data curation DevopsGcpData ModelingAzurePythonEtlAWSAI toolsanalytical workloads
2-4 yrs
Mumbai, India
Skills:
Spark SQLGitAzure Data FactoryPysparkSQL ServerPythonAzure SqlAzure DevOpsCI CD