Search by job, company or skills

Data Engineer

  • Posted 24 minutes ago
  • Be among the first 10 applicants

Job Description

Get to Know Indium Tech

Indium Software is a leading provider of Digital Engineering Services, helping clients drive measurable business

value through technology. We provide services across Application Engineering, Data & Analytics, Cloud

Engineering, Digital Assurance, and Low-Code Development.

Indium is an AI-driven digital engineering company with 5,000+ associates globally and more than 25 years in

business. Its expertise spans Generative AI, Product Engineering, Intelligent Automation, Data & AI, Quality

Engineering, and Gaming.

Indium partners with Fortune 500, Global 2000, and leading technology firms across multiple industries and

geographies.

Learn more: https://www.indium.tech/#

About the Role

We are looking for a Data Engineer to build and maintain scalable, reliable data pipelines and data platforms. The

role involves working across data ingestion, transformation, processing, warehousing, and analytics while

leveraging modern big-data and cloud technologies.

You will work closely with data analysts, data scientists, and engineering teams to ensure high-quality, accessible,

and production-ready data.

Key Responsibilities

• Design, develop, and maintain scalable data pipelines using Python and PySpark.

• Build robust ETL/ELT workflows for ingesting, transforming, validating, and integrating data from multiple sources.

• Write and optimize advanced SQL queries for large-scale data processing and analytics.

• Develop and maintain data warehouse solutions and dimensional data models.

• Work with Hadoop and other big-data technologies to process high-volume datasets.

• Implement data quality, validation, monitoring, and error-handling mechanisms across pipelines.

• Collaborate with cross-functional teams to understand data requirements and translate them into scalable

engineering solutions.

• Optimize data pipelines for performance, reliability, scalability, and cost efficiency.

• Contribute to modern data-platform and AI-enabled data engineering initiatives.

Required Skills

• Strong hands-on experience with Python.

• Strong proficiency in PySpark and distributed data processing.

• Advanced SQL, including complex queries, joins, CTEs, window functions, and query optimization.

• Strong understanding of Data Engineering principles and best practices.

• Hands-on experience with ETL/ELT pipeline development.

• Strong understanding of Data Warehousing concepts and data modeling.

• Experience with Hadoop and Big Data technologies.

• Strong debugging, problem-solving, and communication skills.

Preferred Skills

• Experience with Apache Airflow for workflow orchestration.

• Understanding of Vector Embeddings and their application in modern data and AI systems.

• Exposure to Agentic Frameworks and LLM workflows.

• Familiarity with MCP Servers and modern data platforms.

• Understanding of Semantic Search and retrieval-oriented data systems.

• Experience with cloud platforms such as AWS, GCP, or Azure.

What We Value

• Strong analytical and problem-solving ability.

• Ability to work independently and collaboratively in a fast-paced environment.

• Good understanding of scalable and production-grade data engineering practices.

• Strong communication and stakeholder collaboration skills.

• Curiosity and willingness to learn modern data, cloud, and AI technologies.

Why Join Indium

Work on modern Data Engineering, Big Data, Cloud, and AI-enabled data platform initiatives while contributing

to scalable solutions that create measurable business impact.

More Info

Job Type:
Industry:
Employment Type:

About Company

Job ID: 153617403

Similar Jobs

Hyderabad, India

Skills:

Apache AirflowSparkApache BeamPythonData Platform ManagementData Layer DesigndbtGoogle Cloud EcosystemGoogle Cloud ServicesData Pipeline Development

Hyderabad, India

Skills:

snowflake Data TransformationELTSql QueriesData CleansingMS SQLViewsEtlcomplex joinsdata processing pipelinesstored procedurestemporary tablesData ValidationCTEs

Hyderabad, India

Skills:

Pl SqlDdlInformaticaBodiData IntegrationKinesisOdiRDBMSDatastageData GovernancePythonAWSHadoopScalaEmrSparksqlSSISSqlHiveHiveqlSparkData WarehousingAmazonEtlDataCraftRRedshift SpectrumMDXData LakesCradleGlueKornShell

Hyderabad, India

Skills:

snowflake Sql QueriesPysparkPython ProgrammingDatabricksAWSHCP and HCO data managementAzure blob storageCDW Commercial Data Warehouse

Hyderabad, India

Skills:

tokenization Spark SQLApache SparkGitOcrPythonHugging Face TransformersText ExtractionPDF ParsingVector DatabasesDocument ChunkingData pipelinesCI CD practicesText ProcessingNLP conceptsETL workflowsHTML Parsing

Beware of Scammers

We don’t charge money for job offers