

Search by job, company or skills

Total Experience Required: 7+ years
Mode of Work: Work from Office
Location: Hyderabad
About the Role :
We are looking for a Data Engineer to own and extend the Databricks Lakehouse that powers every Bay6 AI product – Scout, Rebound, Guide, Pathways and Verify. You will work at the center of our data architecture, ingesting client CRM/SIS data (Slate, Banner, EAB Navigate, Stellic, Salesforce, etc.) in the context of complex and potentially fragmented tech stacks inside of Higher Education institutions. You will typically build out a medallion pipeline (bronze/silver/gold) and maintain a multi-tenant Unity Catalog isolation framework across a growing client portfolio. You will partner closely with the engineering team and the AI solutions team to ensure that every agent has clean, governed and well modeled data to work on.
This is a hands-on role in a small, fast-moving environment where you will also be writing pipeline code and making architectural decisions. You must be comfortable working with clients and also be able to articulate decisions or concerns eloquently.
Bay6 AI operates as a headless solution with no native UI. All agent output write into client CRMs. Data accuracy and traceability aren't trivial concerns as they show up directly in a client's production CRM. You will be one of the few engineers who would deeply understand the full data path from a client's SIS all the way to an agent decision, which means real ownership and visibility into product outcomes.
Must Have Skills
Nice to Have Skills
Job Responsibilities
Required Experience:
Educational Qualifications:
Bachelor's/Master's in Computer Science, AI, Engineering or related field.
Job ID: 152468297
Skills:
Azure Data Factory, Pyspark, Sql, Delta Live Tables, Data Bricks
Skills:
snowflake , Hld, Data Migration, Lld, Power Bi, Pl Sql, Tableau, Informatica, Sql, DataStage, DBMS concepts, Data Validation, Custom SQL code
Skills:
Azure Data Factory, Data Modelling, Azure Synapse Analytics, Azure Data Lake, Python, Sql, DataOps, CI CD pipelines, Microsoft Fabric
Skills:
Kafka, Json, ELT, Oracle, Azure DevOps, BigQuery, Apis, Dataproc, Jira, Sql, Git, DB2, Spark, Xml, Data Warehousing, DataFlow, Etl, AlloyDB, CDMP Certified Data Management Professional Certification, Google Pub Sub, LLMs, Ai, Google Cloud Storage, Spanner
Skills:
Apis, Version Control, Maven, Pyspark, Performance Tuning, Apache Spark, Sparksql, Data Modeling, Sql, Jenkins, Databricks, Python, AWS, Scaled Agile methodologies, workflow orchestration, CI CD, Data Fabric, Data Mesh