Search Jobs

Search by job, company or skills

Data Architect

Data Architect

Airtel ATN
Early Applicant
  • Posted 2 months ago
  • Be among the first 10 applicants

Job Description

Location – Gurgaon

Experience – 5-10 years

Role – Big Data Architect.

About the Role:

As a Big Data Engineer, you will play a critical role in integrating multiple data sources, designing scalable data workflows, and collaborating with data architects, scientists, and analysts to develop innovative solutions. You will work with rapidly evolving technologies to achieve strategic business goals.

Must-Have Skills:

  • 4+ year's of mandatory experience with Big data.
  • 4+ year's mandatory experience in Apache Spark.
  • Proficiency in Apache Spark, Hive on Tez, and Hadoop ecosystem components.
  • Strong coding skills in Python & Pyspark.
  • Experience building reusable components or frameworks using Spark
  • Expertise in data ingestion from multiple sources using APIs, HDFS, and NiFi.
  • Solid experience working with structured, unstructured, and semi-structured data formats (Text, JSON, Avro, Parquet, ORC, etc.).
  • Experience with UNIX Bash scripting and databases like Postgres, MySQL and Oracle.
  • Ability to design, develop, and evolve fault-tolerant distributed systems.
  • Strong SQL skills, with expertise in Hive, Impala, Mongo and NoSQL databases.
  • Hands-on with Git and CI/CD tools
  • Experience with streaming data technologies (Kafka, Spark Streaming, Apache Flink, etc.).
  • Proficient with HDFS, or similar data lake technologies
  • Excellent problem-solving skills — you will be evaluated through coding rounds

Key Responsibilities:

  • Must be capable of handling existing or new Apache HDFS cluster having name node, data node & edge node commissioning & decommissioning.
  • Work closely with data architects and analysts to design technical solutions.
  • Integrate and ingest data from multiple source systems into big data environments.
  • Develop end-to-end data transformations and workflows, ensuring logging and recovery mechanisms.
  • Must able to troubleshoot spark job failures.
  • Design and implement batch, real-time, and near-real-time data pipelines.
  • Optimize Big Data transformations using Apache Spark, Hive, and Tez
  • Work with Data Science teams to enhance actionable insights.
  • Ensure seamless data integration and transformation across multiple systems.
  • More Info

    Job Type:
    Industry:
    Function:
    Employment Type:

    Key Skills

    About Company

    Similar Jobs

    7-10 yrs
    Gurugram, India
    Skills:
    Data Modelling, Cloud Architecture, Pyspark, Azure Sql, Sql, Spark Streaming, Azure Synapse Analytics, Azure Data Lake, Apache Kafka, Cosmos DB, Data Governance, Azure Machine Learning, Python, AWS, Distributed computing paradigms, Big Data cloud technologies, Azure OpenAI, Data ingestion programs, ETL methodologies, BI and data analytics databases, Azure Cognitive Services, AI and machine learning workloads
    8-10 yrs
    Gurugram, Gurugram, India
    Skills:
    snowflake , Agile Methodologies, Data Warehousing, Sql, Master data management, Product backlog management, ETL/ELT architecture, Data quality management, Cloud data architectures, Relational Database Design
    12-14 yrs
    Gurugram, Gurugram, India
    Skills:
    snowflake , Data Modelling, Aws Redshift, Pyspark, Sql, Azure Synapse, Databricks, Data Governance, Python, Teradata, GCP BigQuery, Microsoft Fabric
    10-15 yrs
    Noida, India
    Skills:
    Spark, Sql, Clustering, Tensorflow, Pandas, Tableau, Numpy, Power Bi, AWS, Pytorch, Python, Scipy, Git, Machine Learning Algorithms, Redshift, Regression, Classification, Glue, R
    10-18 yrs
    Gurugram, India
    Skills:
    T-sql, SQL Server, Azure Data Factory, Databricks, Python, Etl, Logic App, Azure Data Lake Storage Gen2, Event Hub, Azure Monitoring, Azure Suite Fabric, Microsoft Purview, Stream Analytics, Data Explorer, CosmosDB