

Search by job, company or skills
Job Role - Python Data Engineer
Experience Required - 8+ Years
Mandate Skills - Python, Pyspark and Databricks
Location - Indore or Pune (Hybrid)
Overview:
We are seeking a highly skilled Python and PySpark Data Engineer who is passionate about building robust data solutions from the ground up. The ideal candidate should have a deep understanding of core programming concepts, be highly proficient in Python and PySpark, and possess strong SQL skills. We're looking for someone who thinks out of the box, understands the value of reusable and scalable solutions, and is comfortable working in dynamic, fast-paced environments. Experience with Databricks and Azure Data Factory (ADF) is needed.
Responsibilities:
Design, develop, and maintain Python and PySpark.
Write complex SQL queries for data transformation, validation, and analysis.
Build scalable, reusable code frameworks with a dynamic and modular mindset.
Develop solutions from scratch with clean, well-documented code.
Optimize and refactor existing code for better performance and scalability.
Collaborate closely with data engineers, analysts, and business stakeholders to deliver impactful solutions.
Implement best practices in data engineering, coding standards, and deployment.
Work with Databricks and ADF to orchestrate and manage data workflows.
Requirements:
Strong foundational knowledge and hands-on experience in Python and PySpark.
Proficiency in SQL and experience working with large datasets.
Demonstrated ability to code from scratch, not just maintain or tweak existing codebases.
Ability to think innovatively and approach problems with a solution-oriented mindset.
Understanding of reusable design patterns and dynamic solutioning.
Experience working with Databricks and Azure Data Factory (ADF).
One of the cloud data platforms (Azure, AWS, or GCP) is required
Bachelor's degree in Computer Science, Engineering, or a related field.
About InfoBeans:
InfoBeans is a global digital transformation and product engineering company, enabling businesses to thrive through innovation, agility, and cutting-edge technology solutions. With over 1,700 team members across the globe, we specialize in custom software development, enterprise solutions, cloud, AI/ML, UX, automation, and digital transformation services.
At InfoBeans, we live by our core purpose of Creating WOW!—for our clients, team members, and the community. Our collaborative culture, growth opportunities, and people-first approach make us one of the most trusted and rewarding workplaces.
Link: https://infobeans.ai/
Job ID: 126925989
Skills:
Document Databases Working knowledge of MongoDB, ETL Tools Proficiency in Ab Initio GDE EME Co Operating System, Unix Linux Scripting, SQL PL SQL Querying, Testing QA Exposure to SIT E2E and OAT testing cycles, Messaging Systems Experience with IBM MQ Apache Kafka, Agile Delivery Experience with JIRA, Scheduling Tools Familiarity with Tivoli Workload Scheduler TWS, Security Compliance Awareness of access control vulnerability assessments
Skills:
.NET, data warehouses , Spark SQL, Power Bi, Azure Log Analytics, Powershell Scripting, Hive, Azure Data Factory, Data lakes, Microsoft Azure Data platform, Azure SQL Data Warehouse, Azure Storage Services, Azure Application Insights, Data Bricks, Stream Analytics, data marts, Event Hubs, Azure SQL DB, Azure Analysis Services
Skills:
Java, Unix, Apache Flink, Data Modeling, Schema Design, Pyspark, Data Cleansing, Apache Spark, Data Warehouse Concepts, Shell Scripting, Sql, ELT, Apache Airflow, Linux, Apache Kafka, Restful Apis, Python, Etl, Data Transformation, Data Quality Validation
Skills:
snowflake , Kafka, ELT, Devops, Spark, Incident Management, Databricks, Restful Apis, Azure, Etl, AWS, Airflow, Monitoring
Skills:
Data Governance, Data Modeling, Sql, Python, data orchestration workflows, data quality frameworks, distributed data processing frameworks, ETL pipelines