
Search by job, company or skills
About Us
We are an early-state AI startup building a structured research data platform for the investment research domain. We serve professional investment teams focused on industry fundamental research. With our first batch of seed clients (mainstream hedge funds) already onboarded, we are a lean and highly efficient team.
We are looking for an early core engineer to work alongside the founding team to build this data platform from 0 to 1: from data ingestion and structured parsing, to query services for analysts. This is an end-to-end role where you will have complete ownership.
Core Responsibilities
1. Data Ingestion
2. Document Parsing
3. Data Warehouse Design
4. Pipeline Orchestration & Monitoring
5. LLM Application Engineering
6. Backend Service Layer
7. Cloud Infrastructure
Key Requirements
1. Engineering Experience
2. Technical Stack
3. Document Processing
4. AI & LLM Expertise
5. Data Architecture
6. Soft Skills & Communication
7. Strong Preferred
Preferred Qualifications (Bonus Points)
What We Offer
Job ID: 153602583
Skills:
Java, Data Modelling, Pyspark, Kafka, Sql, ELT, Docker, Terraform, Kubernetes, Python, Etl, Airflow, cdc, Flink, CI CD, Iceberg, lakehouse architecture, NiFi, Cloudera Machine Learning, observability tooling
Skills:
Apache Spark, Data Integration, Data Modelling, Data Transformation, Etl Development, elastic search/open search, python/scala
Skills:
snowflake , Hadoop, Microsoft Power Bi, Tableau, Data Integration, Google Cloud, Hive, Hue, Data Architecture, Data Lake, Databricks, Azure, Talend, Python, AWS, Etl, Alibaba Cloud
Skills:
data warehouses , snowflake , Java, Data Warehouse Concepts, Scala, Dimensional Modeling, Sql, Data Lake, Python, Airflow, dbt-core, Big Data processing frameworks, SCDs, lake house, data mesh, modern data warehouse
Skills:
BigQuery, Aws Redshift, Hadoop, Pyspark, Kafka, AWS Athena, Sql, Nosql, Gcp, MySQL, Spark, Sns, Data Lake, Python, AWS, Aws S3, Etl, Apache Iceberg, Parquet, Airflow, AWS MSK, Firehose