Search by job, company or skills

Data Engineer

Data Engineer

Bahwan Cybertek
5-7 Years
Not Disclosed
Early Applicant
  • Posted 15 hours ago
  • Be among the first 10 applicants

Job Description


Required Skills & Experience

Around 5 years of professional experience in Data Engineering, Python development, or a related role.

Strong hands-on expertise in Python, particularly for:

Data wrangling and transformation

ETL/ELT development

File and data processing

Automation and scripting

Data validation and cleansing

Strong knowledge of at least one relational SQL database such as:

PostgreSQL

MySQL

Oracle

SQL Server

Strong understanding of SQL, including joins, subqueries, CTEs, window functions, aggregations, and query optimization.

Practical experience with AWS cloud services used in data engineering.

Experience developing RESTful APIs using FastAPI.

Good understanding of API concepts including HTTP methods, request/response handling, authentication, status codes, error handling, and API integration.

Experience with Git/version control and standard software development practices.

Good understanding of data engineering concepts, including data pipelines, ETL architecture, data quality, and data integration.

Good to Have

Hands-on experience with PySpark and distributed data processing.

Experience with NoSQL databases such as DynamoDB, or Cassandra.

Experience with AWS services such as S3, Lambda, Glue, Athena, Redshift, EMR, RDS, Step Functions, EventBridge, and CloudWatch.

Experience with workflow orchestration tools such as Apache Airflow or AWS Step Functions.

Experience with Docker and containerized applications.

Familiarity with CI/CD pipelines and DevOps practices.

Knowledge of data warehousing and dimensional data modeling.

Experience working with large-scale datasets and performance optimization.

Familiarity with cloud-based data lake and data warehouse architectures.

Technical Skills

Skill Area Expected Knowledge

Programming Python – Advanced

Data Engineering ETL/ELT, data wrangling, transformation, validation

Database SQL, Relational Databases

Cloud AWS

API Development FastAPI, REST APIs

Big Data PySpark – Good to Have

NoSQL MongoDB / DynamoDB / Cassandra – Good to Have

Data Formats JSON, CSV, XML

Version Control Git

DevOps CI/CD, Docker – Good to Have

Skills

Programming Python – Advanced

Data Engineering ETL/ELT, data wrangling, transformation, validation

Database SQL, Relational Databases

Cloud AWS

API Development FastAPI, REST APIs

Big Data PySpark – Good to Have

NoSQL MongoDB / DynamoDB / Cassandra – Good to Have

Data Formats JSON, CSV, XML

Version Control Git

DevOps CI/CD, Docker – Good to Have

More Info

Job Type:
Industry:
Employment Type:

About Company

Similar Jobs

3-10 yrs
Hyderabad, India
Skills:
S3, Aws Services, Scala, Apache Spark, Emr, Redshift, Lambda, Kinesis, Cloudwatch, Iam, Glue, Athena
5-10 yrs
Bengaluru, India
Skills:
Pyspark, Spark, Azure Databricks, Sql, Azure DevOps, Delta Lake, Event Hub
5-10 yrs
Noida, India
Skills:
S3, Github, Pyspark, Lambda, Ec2, Kinesis, Docker, Shell scripting, Python, AWS, Scala, Redshift, Sql, Git, Kubernetes, GenAI, Airflow, AI coding assistants, LLM tools, RAG pipelines, LangChain, prompt engineering, dbt, Glue, LlamaIndex
6-8 yrs
Gurugram, Gurugram, India
Skills:
Github, Pyspark, Pandas, Numpy, Gcp, Docker, Databricks, Azure, Python, AWS, scikit-learn, Great Expectations, Kedro
3-5 yrs
Hyderabad, India
Skills:
Google Cloud Platform (GCP), BigQuery, Schema Design, Data Modeling, Sql, ELT, DataFlow, Python, Etl, Fine-tuning of models, Vector databases, Generative AI technologies, BigQuery ML, Vertex AI, RAG, LLM integrations, Data pipelines, Pub/Sub