Search by job, company or skills

Senior Data Engineer

Senior Data Engineer

Recro
5-7 Years
Not Disclosed
  • Posted 17 hours ago
  • Be among the first 10 applicants

Job Description

Data Engineer JD

Location - BLR (5 days WFO)

EXP - 5+ years

Job Role:

Develop long-term vision for a highly scalable data platform, data management and Data Ops practices.

Design and architect data flows, data management in Hadoop or Cloud environment which are scalable, repeatable and eliminate time consuming steps Promote Data Ops approach to automate the provision of data, testing and monitoring and to shorten development cycles and increase deployment frequency.

Establish development and data governance processes to build mature data pipelines, CI/CD, test coverages, etc.

Evaluate, provide insights and recommendations on tools and technology strategy for analytics data platforms and applications in conjunction with Enterprise Architecture team

Ability to lead data engineering workstreams with a product mindset

Who are we looking for

Bachelor's or master's degree in computer science, Information Systems or equivalent field.

At least 5+ years of experience in building data flows and data management on modern big data tech stack Data Strategy: Understands, articulates, and applies principles of the defined strategy to routine business problems that involve a single function.

Data Transformation and Integration: Extracts data from identified databases. Creates data pipelines and transform data to a structure that is relevant to the problem by selecting appropriate techniques.

Develops knowledge of current analytics trends.

Data Source Identification: Supports the understanding of the priority order of requirements and service level agreements.

Helps identify the most suitable source for data that is fit for purpose.

Demonstrates expertise in writing complex, highly optimized queries across large data sets Strong experience in using ETL framework (eg. Airflow, Oozie, Jenkins etc.) to build and deploy production-quality ETL pipelines.

Experience in ingesting and transforming structured and unstructured data from internal and third-party sources into dimensional models.

Knowledge of data structures and distributed computing.

Should be comfortable in manipulation and analysis of high-volume data from variety of internal and third-party sources.

Experience in one or more programming languages like Python or PySpark and moderate knowledge on unix scripting.

More Info

Job Type:
Industry:
Employment Type:

Key Skills

Cloud environment

Data Ops

CI/CD

Data pipelines

Data transformation and integration

About Company

Similar Jobs

3-5 yrs
Bengaluru, India
Skills:
BigQuery, Google Cloud Platform, Pyspark, Spark, Dataproc, Sql, Airflow
6-9 yrs
Bengaluru, India
Skills:
Java, S3, Scala, AWS Glue, Emr, Redshift, Sql, Big Data Technologies, Nosql, Lambda, Python, AWS, data quality governance, data pipeline troubleshooting, security best practices
4-7 yrs
Bengaluru, India
Skills:
table design , snowflake , BigQuery, Data Modeling, Redshift, Sql, ELT, Jenkins, Query Tuning, Python, Etl, Airflow, ML feature engineering, GitHub Actions, CI/CD pipelines, dbt, AWS SageMaker, Dagster, workload management
7-9 yrs
Bengaluru, India
Skills:
Hadoop, Etl Development, Scala, Data Modeling, Dataproc, Hive, Presto, Spark, Big Data Technologies, DataFlow, Azure, Python, AWS, Big Query, Airflow, Data lake architecture, Cloud Composer, GCS, GCP technologies
5-8 yrs
Bengaluru, India
Skills:
Pyspark, Azure Databricks, Sql, Git, Data Quality Governance, ADLS Gen2, Medallion Architecture, DLT, Auto Loader, Unity Catalog, Delta Lake, Structured Streaming, CI-CD, Databricks Workflows