Search by job, company or skills

Senior Data Engineer

Early Applicant
  • Posted 22 hours ago
  • Be among the first 10 applicants

Job Description

At Nat Habit, we are building efficacious personal care using natural ingredients while doing primary research to build formulations from a first principle basis to find cures, study the underlying impact on different markers and then conduct clinical studies to prove efficacy of products. Check us out on www.nathabit.in

The founding team has a strong startup experience and is well funded and backed by top angel investors and tier 1 institutional investors.

Job Summary: Nat Habit is looking for a Data Engineer who enjoys building production-grade data systems. Youwill be responsible for designing, developing, deploying, and maintaining reliable data pipelines thatsupport business analytics and intelligence.

This role is ideal for engineers who enjoy solving real engineering problems, writing clean code,and building scalable data platforms.

Key Responsibilities:

  1. Design, build and maintain scalable ETL/ELT pipelines using SQL, Python and tools such as Dagster, Airflow or dbt to ingest, transform and load data into data platforms.
  2. Develop scalable and reliable data ingestion workflows.
  3. Build incremental data pipelines capable of handling millions of records efficiently.
  4. Design reporting datasets and analytical data models.
  5. Work extensively with ClickHouse or other OLAP database for analytical workloads.
  6. Optimize query performance, data storage and processing efficiency through advanced SQL techniques, partitioning, indexing strategies and workload optimization.
  7. Monitor pipeline execution, troubleshoot failures, and improve system reliability.
  8. Solid understanding of batch and streaming data processing techniques.
  9. Define and enforce data governance policies including data security, access control, masking and lifecycle management within data platforms.

Ideal candidate will have:

  • Strong proficiency in Python, including Pandas, NumPy, PySpark.
  • Hands on experience with ETL / ELT and data integration tools, such as Apache Airflow, Pentaho, or Dagster
  • One or more years of experience with RDBMS and OLAP databases like Clickhouse, Redshift.
  • Good working knowledge of containers like Docker or Podman
  • Experience working with IaaC like terraform
  • Experience in data warehousing & data archival.
  • Should have good understanding & knowledge about observability.
  • Hands-on experience deploying projects beyond local development.
  • Experience in version control tool: Git

Do not apply if following is applicable to you:

  • Backend engineers who only enjoy request/response APIs and CRUD, and don't want to own end-to-end data pipelines, scheduling, and reliability.
  • Engineers who do not have experience with business intelligence and related tooling.
  • Data analysts who mainly write ad-hoc SQL and build dashboards, and do not have engineering experience eg: version control, testing, CI/CD, and data infrastructure.
  • People who don't naturally question data quality – if obviously wrong or inconsistent metricsdon't bother you, this won't be a good fit.

Work

Location: Udyog Vihar - Sector 18, Gurgaon

www.nathabit.in

www.instagram.com/nathabit.in

Working Days: Mon - Sat (2nd/4th Sat are off)

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 151635307

Similar Jobs

Noida, India

Skills:

data engineering snowflake GithubPysparkApache SparkSqlBitbucketApache KafkaGitlabDatabricksPythonDelta Lake on DatabricksData Quality Validation

Noida, India

Skills:

data warehouses snowflake Power BiData ModelingData LineageJiraSqlDimensional ModelingData VisualizationPythonAWSELT patternsSemantic layersData ValidationAnalytics engineeringData automationProject ManagementVersion-controlled analytics code

Noida, India

Skills:

JavaCassandraScalaPostgreSQLApache SparkKafkaSpring BootSqlApache NifiRESTGcpDockerMongoDBOracleKubernetesPythonAWSAirflowFlinkGRPC

Gurugram, Gurugram, India

Skills:

data engineering Machine LearningArtificial IntelligenceApache SparkSqlData ScienceGitDockerApache KafkaPythonAWSEtlLangChainGenerative AILLMsLlamaIndex

Gurugram, India

Skills:

Data ModelingSqlSparkData IntegrationPythondata pipeline developmentOneLakeschema managementdata quality frameworksReconciliationincremental processingPipelinesNotebooksperformance optimizationpartitioningLakehouse architectureMicrosoft FabricAzure data servicesValidation