Job Overview
- Role: Data Engineer (Snowflake & GenAI focus)
- Experience: 4+ Years
- Location: Pune
About Us & Team Scale
Helpshift is on a mission to modernize customer service through mobile-first, in-app chat, AI-powered chatbots, and automation for global brands like Supercell, EA, Brex, and Square. The Data Platform team manages the infrastructure powering in-app analytics and data integrations, handling over 2 million events per minute and processing 1+ TB of data daily across 900M+ monthly active users and 1000 peak VMs.
Key Responsibilities
- Cortex & Agentic Architecture: Architect and implement advanced Agentic use cases leveraging Snowflake Cortex, embedding LLM capabilities (summarization, anomaly detection, automated insights) into core data workflows.
- Pipeline Development: Design, build, and deploy highly scalable data pipelines natively within Snowflake using Snowpark and Python/SQL Stored Procedures.
- Cost & Performance Optimization: Rigorously monitor, troubleshoot, and optimize Snowflake compute costs and query performance, specifically focusing on the execution efficiency of LLMs and Agentic functions.
- Data Modeling & Architecture: Apply expert data modeling practices to design clean, scalable, and secure data schemas supporting traditional analytics and AI/ML workloads.
- Rapid Delivery: Operate in a fast-paced environment, owning end-to-end technical delivery from architecture to production deployment.
- Engineering Standards: Maintain high standards through technical design documents, code reviews, and CI/CD best practices for Snowflake artifacts.
Requirements
- Experience: 4+ years of Data Engineering experience building complex, high-performance Analytics and AI-driven production systems.
- Snowflake Expertise: Extensive, hands-on experience with Snowflake, including advanced data modeling, schema design, and native Snowflake features.
- Snowflake Cortex & GenAI: Demonstrable experience leveraging Snowflake Cortex AI to build scalable GenAI and Agentic use cases directly within the data platform.
- Core Tech Stack: Expert-level proficiency in Python and SQL, with strong foundational knowledge of data structures, algorithms, and software engineering principles.
- LLM & Cost Optimization: Deep understanding of performance tuning and cost optimization tailored for LLMs, vector searches, and Agentic systems within cloud data platforms.
- Cloud & Operations: Familiarity with modern cloud platforms (AWS/GCP/Azure) integrated with Snowflake, alongside solid production ops knowledge (observability, data quality, reliability, RBAC).
- Education & Communication: Bachelor's Degree in Computer Science, Engineering, or related field (or equivalent practical experience), paired with strong verbal and written communication skills.