Search by job, company or skills

Remote Data Engineer

Remote Data Engineer

Turing
3-5 Years
Not Disclosed
Early Applicant
  • Posted 21 days ago
  • Be among the first 30 applicants

Job Description

About Turing:

Based in San Francisco, California, Turing is the world's leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports customers in two ways: first, by accelerating frontier research with high-quality data, advanced training pipelines, plus top AI researchers who specialize in coding, reasoning, STEM, multilinguality, multimodality, and agents; and second, by applying that expertise to help enterprises transform AI from proof of concept into proprietary intelligence with systems that perform reliably, deliver measurable impact, and drive lasting results on the P&L

Role Overview:

We are looking for experienced Software Engineers (SWE Bench – Data Engineer / Data Science) to contribute to benchmark-driven evaluation projects focused on real-world data engineering and data science workflows. This role involves hands-on work with production-like datasets, data pipelines, and data science tasks to help evaluate and improve the performance of advanced AI systems.

The ideal candidate has strong foundations in data engineering and data science, with the ability to work across data preparation, analysis, and model-related workflows in real-world codebases.

What does day-to-day life look like

  • Work with structured and unstructured datasets to support SWE Bench-style evaluation tasks.
  • Design, build, and validate data pipelines used in benchmarking and evaluation workflows.
  • Perform data processing, analysis, feature preparation, and validation for data science use cases.
  • Write, run, and modify Python code to process data and support experiments locally.
  • Evaluate data quality, transformations, and outputs for correctness and reproducibility.
  • Create clean, well-documented, and reusable data workflows suitable for benchmarking.
  • Participate in code reviews to ensure high standards of code quality and maintainability.
  • Collaborate with researchers and engineers to design challenging, real-world data engineering and data science tasks for AI systems.

Requirements:

  • Minimum 3+ years of overall experience as a Data Engineer, Data Scientist, or Software Engineer (data-focused).
  • Strong proficiency in Python for data engineering and data science workflows.
  • Demonstrable experience with data processing, analysis, and model-related workflows.
  • Solid understanding of machine learning and data science fundamentals.
  • Experience working with structured and unstructured data.
  • Ability to understand, navigate, and modify complex, real-world codebases.
  • Experience writing readable, reusable, maintainable, and well-documented code.
  • Strong problem-solving skills, including experience with algorithmic or data-intensive problems.
  • Excellent spoken and written English communication skills.

Perks of Freelancing With Turing:

  • Work in a fully remote environment.
  • Opportunity to work on cutting-edge AI projects with leading LLM companies.

Offer Details:

  • Commitments Required: At least 4 hours per day and minimum 20 hours per week with overlap of 4 hours with PST.
  • Engagement Type: Contractor assignment (no medical/paid leave)
  • Duration of Contract: 3 months (adjustable based on engagement)

After applying, you will receive an email with a login link. Please use that link to access the portal and complete your profile.

Know amazing talent Refer them at turing.com/referrals, and earn money from your network.

More Info

Job Type:
Industry:
Employment Type:

Key Skills

About Company

Similar Jobs

5-8 yrs
Bengaluru, India
Skills:
BigQuery, Pyspark, Dataproc, Data Modeling, Sql, ELT, Git, Python, Etl, GCP data services, Cloud Dataflow, Pub Sub, Cloud Composer, CI CD tools, performance optimization, partitioning, GCS
6-8 yrs
Bengaluru, India
Skills:
Azure Data Factory (ADF), T-sql, Data Warehouse Concepts, Pyspark, SQL Server, Sql, Git, Python, Star Schema, Azure DevOps, Generative AI, AI-Assisted Development, Claude Claude Code, Fact Dimension modeling, Azure SQL Database, Stored Procedures, CI/CD, Microsoft Fabric
6-10 yrs
Bengaluru, India
Skills:
Agile Methodologies, Pyspark, Sql, Big Data Technologies, Microservices, Git, Gcp, Apache Kafka, Spark, Rest Apis, Azure, Python, AWS, Kafka Connect, CI CD tools, ETL ELT pipelines, Kafka Streams
7-11 yrs
Bengaluru, India
Skills:
PostgreSQL, Sql, ELT, Python, Etl, canonical data modeling, orchestration tools, data contracts, AI coding assistants, lakehouse concepts, data-quality validation, cloud data platforms, config-driven architecture
5-8 yrs
Bengaluru, India
Skills:
Pyspark, Azure Databricks, Sql, Git, Data Quality Governance, ADLS Gen2, Medallion Architecture, DLT, Auto Loader, Unity Catalog, Delta Lake, Structured Streaming, CI-CD, Databricks Workflows