Search by job, company or skills

Platform Engineer

Platform Engineer

Tata Electronics Tepl
1-4 Years
Not Disclosed
Quick Apply
  • Posted 27 days ago
  • Over 100 applicants have applied

Job Description

Job Responsibilities -

  • Architect and implement a scalable, offline Data Lake for structured, semi-structured, and unstructured data in an on-premises, air-gapped environment.
  • Collaborate with Data Engineers, Factory IT, and Edge Device teams to enable seamless data ingestion and retrieval across the platform.
  • Integrate with upstream systems like MES, SCADA, and process tools to capture high-frequency manufacturing data efficiently.
  • Monitor and maintain system health, including compute resources, storage arrays, disk I/O, memory usage, and network throughput.
  • Optimize Data Lake performance via partitioning, deduplication, compression (Parquet/ORC), and implementing effective indexing strategies.
  • Select, integrate, and maintain tools like Apache Hadoop, Spark, Hive, HBase, and custom ETL pipelines suitable for offline deployment.
  • Build custom ETL workflows for bulk and incremental data ingestion using Python, Spark, and shell scripting.
  • Implement data governance policies covering access control, retention periods, and archival procedures with security and compliance in mind.
  • Establish and test backup, failover, and disaster recovery protocols specifically designed for offline environments.
  • Document architecture designs, optimization routines, job schedules, and standard operating procedures (SOPs) for platform maintenance.
  • Conduct root cause analysis for hardware failures, system outages, or data integrity issues.
  • Drive system scalability planning for multi-fab or multi-site future expansions.

Essential Attributes (Tech-Stacks) -

  • Hands-on experience designing and maintaining offline or air-gapped Data Lake environments.
  • Deep understanding of Hadoop ecosystem tools: HDFS, Hive, Map-Reduce, HBase, YARN, zookeeper and Spark.
  • Expertise in custom ETL design, large-scale batch and stream data ingestion.
  • Strong scripting and automation capabilities using Bash and Python.
  • Familiarity with data compression formats (ORC, Parquet) and ingestion frameworks (e.g., Flume).
  • Working knowledge of message queues such as Kafka or RabbitMQ, with focus on integration logic.
  • Proven experience in system performance tuning, storage efficiency, and resource optimization.

Qualifications -

  • BE/ ME in Computer science, Machine Learning, Electronics Engineering, Applied mathematics, Statistics.

Desired Experience Level -

  • 4 Years relevant experience post Bachelors
  • 2 Years relevant experience post Masters
  • Experience with semiconductor industry is a plus

More Info

Job Type:
Industry:
Function:
Employment Type:

Key Skills

About Company

Tata Electronicsis a prominent global player in the electronics manufacturing industry, with fast-emerging capabilities in Electronics Manufacturing Services, Semiconductor Assembly and Test, Semiconductor Foundry, and Design Services. Established in 2020 as a greenfield venture of the Tata Group, the company aims to serve global customers through integrated offerings across a trusted electronics and semiconductor value chain.

Similar Jobs

Bengaluru, India, Remote
Skills:
change data capture , Java, Amazon Web Services, Kotlin, Python, Kubernetes, Spark, Multi region data architectures, Pinot, Batch and streaming architectures, ClickHouse, Druid, Delta Lake, dbt, Event driven ingestion, Iceberg, StarRocks
6-8 yrs
Ahmedabad, India
Skills:
Git, Restful Apis, Python, AWS, Microservices, Event-Driven Architecture, AWS Boto3, Serverless Architecture, NoSQL Databases
4-9 yrs
Ahmedabad
Skills:
Python, Data Lake, Hadoop, HBase, Yarn, Zookeeper
Ahmedabad, India
Skills:
Object-Oriented Programming (OOP), Aws Services, Solid Principles, Git, Docker, Restful Apis, Python, Kubernetes, Event-driven architecture design, NoSQL databases, Serverless development, AWS Boto3 SDK
5-7 yrs
Ahmedabad, India
Skills:
MLops, Terraform, Python, LangChain, LLMOps, Anthropic APIs, LlamaIndex, API gateways, Observability tools, Go or Java