Search Jobs

Search by job, company or skills

Senior Data Scientist

Senior Data Scientist

Akaike Technologies
Early Applicant
  • Posted 5 days ago
  • Be among the first 20 applicants

Job Description

Senior Data Scientist

Experience: 4+ Years | Location: Bengaluru (Hybrid) |Team: Data Science & AI

About the Role

Akaike Technologies builds agentic AI and machine-learning systems that power decision-making for global enterprises. We want a Senior Data Scientist who understands what makes agentic systems reliable and can architect complex, high-accuracy, human-in-the-loop AI for demanding clients. Roughly 70% Generative AI and agents, 30% classical ML and large-scale data, with room to move across problems as priorities shift.

Key Responsibilities

Generative AI & Agentic Systems (primary focus)

  • Agentic Systems: Design and ship production multi-agent systems (LangGraph/CrewAI; ReACT, Agent-Critique) that generate reliable, structured, high-accuracy outputs.
  • Grounding & Reliability: Engineer grounding, source traceability, and hallucination control so outputs hold up in production and regulated settings.
  • Evaluation: Build quantitative evaluation for generation quality — faithfulness, hallucination, consistency, and LLM-as-a-judge frameworks.
  • Human-in-the-Loop: Design configurable, human-in-the-loop workflows (review gates, co-pilot editing) that scale across clients and use cases.
  • Retrieval & Tuning: Build production RAG and agentic retrieval; apply PEFT/LoRA on open-source models (Llama 3, Mistral) where it beats prompting.

Classical ML & Data at Scale (secondary focus)

  • Custom Modeling: Build bespoke models for targeting, budget optimization, and churn, including PU and single-class learning on sparse, noisy data.
  • Applied Deep Learning: Apply 1D/2D CNNs, LSTMs, and embeddings to non-text data (time-series, behavioral logs) where they add real signal.
  • Big Data: Write optimized PySpark/SparkSQL on Databricks over billions of rows, with feature stores consistent across training and inference.

Architecture, MLOps & Delivery

  • System Design: Architect end-to-end systems with deliberate latency/cost/accuracy trade-offs and modular, reusable components.
  • Measurement: Bring statistical rigor — offline and human eval, A/B where it fits, drift detection, and automated retraining.
  • Deployment: Ship scalable pipelines on AWS (Bedrock, Lambda, Step Functions) and FastAPI.

Leadership & Stakeholders

  • Mentorship & Clients: Mentor juniors, run rigorous code reviews, and serve as the technical point of contact for clients — explaining model limitations and risk without overselling.

Must-Have Skills

  • Advanced GenAI/agent frameworks (LangChain, LlamaIndex, LangGraph, DSPy), with at least one agentic system or complex RAG pipeline shipped to production.
  • PyTorch or TensorFlow; solid grasp of attention, encoder-decoder architectures, and embeddings.
  • Expert PySpark and SQL; comfortable debugging and optimizing Spark on Databricks.
  • Python (OOP, typing, rigorous standards) and a strong experimental / statistical mindset.
  • Clear communicator who can explain complex AI — and its risks — to business leaders.

Nice to Have

  • Pharma, life-sciences, or other regulated-domain background.
  • Text-to-SQL, GraphRAG (Neo4j), or serving open-source models (vLLM/TGI).
  • Custom loss functions / attention modifications, or open-source and publication contributions.

Benefits & Perks

  • Competitive ESOP grants.
  • Working with Fortune 500 companies and world-class teams.
  • Publishing papers and attending conferences.
  • Networking events, conferences, and seminars.
  • Visibility across all functions at Akaike — sales, pre-sales, lead generation, marketing, and hiring.

More Info

Job Type:
Industry:
Function:
Employment Type:

Key Skills

LangChain

embeddings

LSTMs

Generative AI

Agentic Systems

LangGraph

DSPy

2D CNNs

LlamaIndex

About Company

Similar Jobs

4-7 yrs
Bengaluru, India
Skills:
causal inference , Databricks, Code Review, Python, double debiased ML, uncertainty quantification, heterogeneous treatment effects, scikit-learn, conformal prediction, regularized regression, gradient boosting, DoWhy, Experimental Design, EconML, optimization logic, Git-based workflows
5-7 yrs
Bengaluru, India
Skills:
Tensorflow, SAP, Pytorch, Erp, Python, Deep Learning, Time Series Forecasting, Scheduling, Optimization, Mes
6-8 yrs
Bengaluru, India
Skills:
Gcp, Pyspark, Sklearn, Azure, Tensorflow, Python, Pytorch
3-5 yrs
Bengaluru, India
Skills:
Tensorflow, Pytorch, Etl Tools, Python, Machine Learning Algorithms, Hadoop, Neural Nets, Hive, Spark, Keras, Clustering, anomaly detection, machine learning techniques, demand sensing, Heuristic LP, R, Market Intelligence, stochastic models, Map Reduce, SQL databases, GA, Gurobi, ensemble learning, time series optimizations, Regression, scalable ML frameworks, forecasting optimization, feature engineering
5-7 yrs
Bengaluru, India
Skills:
Machine Learning, Pyspark, Databricks, Python, Sql, Deep Learning, Ai, Data Platforms