Title: Senior AI Engineer – Vision Language Models (VLM) / LLM – Medical Imaging
Experience: 8+ Years
Location: Bengaluru, Hybrid (Collaborating with the Netherlands & India teams)
Budget: 50LPA
Job type: 1 year contract
About the Role
We are looking for a Senior AI Engineer with deep expertise in Large Language Models (LLMs), Vision Language Models (VLMs), and Medical Imaging AI to help build next-generation multimodal healthcare solutions.
You will play a key technical role in designing, fine-tuning, evaluating, and deploying AI models capable of understanding medical images and generating clinically meaningful reports. This position requires hands-on expertise, architectural ownership, and the ability to work independently across distributed teams while maintaining high standards of quality, safety, and technical excellence.
Key Responsibilities
- Design and fine-tune multimodal AI models for medical image understanding and report generation.
- Develop and optimize Vision Language Models (VLMs) and Large Language Models (LLMs) for healthcare applications.
- Build scalable training pipelines using PyTorch, MONAI, and Hugging Face.
- Fine-tune foundation models using LoRA, QLoRA, and other parameter-efficient techniques.
- Develop DICOM-based preprocessing pipelines for Chest X-ray, Neuro CT, and MSK X-ray datasets.
- Perform model evaluation, error analysis, uncertainty estimation, and calibration.
- Improve model robustness through dataset engineering, augmentation, and class imbalance handling.
- Evaluate report generation using metrics such as BLEU, ROUGE, CheXbert, and RadGraph.
- Optimize distributed GPU training and cloud-based AI workflows.
- Collaborate with engineering and research teams across Europe and India.
- Provide technical leadership and proactively identify architectural and delivery risks.
Required Skills
- 10+ years of experience in AI, Machine Learning, or Computer Vision.
- Strong experience with Large Language Models (LLMs).
- Hands-on experience with Vision Language Models (VLMs).
- Expertise in PyTorch, MONAI, and Hugging Face.
- Experience fine-tuning multimodal foundation models using LoRA or QLoRA.
- Strong medical imaging AI experience involving Chest X-ray, Neuro CT, or MSK X-ray.
- Experience with DICOM data pipelines.
- Knowledge of image preprocessing, augmentation, dataset quality, and model evaluation.
- Experience with distributed training and GPU optimization.
- AWS cloud experience.
- Excellent communication and stakeholder management skills.
- Ability to work independently across global teams.
Preferred Experience
Experience with one or more of the following:
- MedGemma
- LLaVA-Med
- BioViL-T
- GLoRIA
- CheXzero
- MedSigLib
- Vision-language alignment
- Clinical report generation
- Healthcare AI or Radiology AI