Search Jobs

Search by job, company or skills

Generative AI Engineer

Generative AI Engineer

Cerebry
  • Posted 2 months ago
  • Over 50 applicants have applied

Job Description

About the Role:-

We are seeking a Generative AI Engineer with 2+ years of experience to build and scale production-ready Agentic Systems using Large Language Models (LLMs). You will work on RAG pipelines, agent workflows, evaluation, and deployment of reliable AI systems.

Key Responsibilities:-

-Build retrieval-augmented generation (RAG) pipelines with embeddings, hybrid search, and reranking

-Implement agent orchestration, tool/function calling, and prompt management

-Develop evaluation, monitoring, and observability for LLM systems

-Ensure AI safety, governance, and data privacy best practices

-Optimize performance using caching, batching, and streaming

-Package solutions as APIs/SDKs and deploy using cloud-native tools

Tech Stack:-

-Languages: Python (primary), TypeScript

-Frameworks: FastAPI/Flask, LangChain/LlamaIndex

-LLMs: OpenAI, Anthropic, Google, Azure OpenAI, AWS Bedrock

-Infra: Docker, CI/CD, Terraform/CDK

-Cloud: AWS / Azure / GCP (any one)

Requirements:-

-2+ years of software or AI engineering experience

-Strong Python and backend development skills

-Hands-on experience with LLMs and cloud deployments

-Understanding of scalable systems and data security

What We Offer:-

-Work on cutting-edge Generative AI products

-High ownership and learning opportunities

-Competitive compensation and growth

More Info

Job Type:
Industry:
Employment Type:

Key Skills

OpenAI

LlamaIndex

Azure OpenAI

LangChain

AWS Bedrock

Anthropic

About Company

Similar Jobs

3-5 yrs
Noida, India
Skills:
PytorchPythonLangChainvector searchRAG embeddingsfine-tuningHugging Facemodel deploymentvLLMprompt engineeringLangGraphOllamaQwenMistraldocument indexinginference optimizationTGILlamaOpenAIquantizationLlamaIndex
5-7 yrs
Noida, India
Skills:
Azure)Cloud platforms (AWSLarge Language Models (LLMs)GcpPythonApi IntegrationMLopsSystem DesignPrompt engineeringModel monitoringRAG systemsHugging FaceClaudeAzure OpenAIEmbeddingsEvaluation metricsGenerative AIOpenAI GPTVector databasesSemantic search technologiesAnthropicFine-tuning techniquesDeployment pipelinesOpenAI
4-8 yrs
Gurugram, India
Skills:
TensorflowNlpPytorchDockerPythonAWSGcpAzureKubernetesGloVeHugging FaceMLflowPineconeMLOps practicesLLM architectureWord2VecRAG pipelinesChromaDBfine-tuning techniquesAPIs from major AI providersLangChainLLM training methodologiesLoRAQLoRAsentence transformerscloud computing platforms
3-5 yrs
Noida, India
Skills:
PytorchLinuxDockerPython ProgrammingPEFTHugging Face Transformersopen-source LLMsmodel routingprompt engineeringGPU-based inference and optimizationEvaluationstructured extractionLoRANVIDIA GPU infrastructurererankingRAGGenerative AI LLMs
5-7 yrs
Gurugram, India, Gurugram
Skills:
Deep LearningMachine LearningAWSCloud ServicesPythonAzureGcpAutoGenNLP LibrariesLang ChainDistributed Computing FrameworksLang GraphOpenAI Assistants API