Search Jobs

Search by job, company or skills

AI Safety & Evaluation Lead

AI Safety & Evaluation Lead

the lhr group
  • Posted 7 hours ago
  • Be among the first 10 applicants

Job Description

Role Overview: As the AI Safety & Evaluation Lead, you will ensure AI systems are accurate, safe, unbiased, secure, and compliant. With students as the primary users, you will focus on student psychology, content safety, and handling sensitive queries in an age-appropriate manner. This is a 0-to-1 role requiring high ownership, hands-on execution, adaptability, and comfort with ambiguity and fast-paced work.

Key Responsibilities

  • Own AI evaluation frameworks, including automated, human-in-the-loop, and benchmark testing.
  • Lead red-teaming, adversarial testing, and vulnerability identification across AI models.
  • Define AI safety, responsible AI, bias, fairness, and child-safe content standards.
  • Build and maintain evaluation tools using platforms such as RAGAS, LangSmith, PromptBench, and LMEval.
  • Conduct bias, fairness, hallucination, grounding, and response-quality assessments across languages and regions.
  • Manage AI risk assessments and translate findings into actionable model or prompt improvements.
  • Establish AI governance documentation, audit trails, and safety reports.
  • Embed evaluation and safety checks into AI release pipelines with engineering teams.
  • Track relevant AI governance and regulatory frameworks, including NIST AI RMF and India/EU AI regulations.

Requirements

  • Strong experience in AI/LLM safety, evaluation, and risk assessment.
  • Expertise in model evaluation, red-teaming, and adversarial testing.
  • Strong understanding of bias, fairness, hallucination, grounding, and content safety.
  • Experience working with production AI/LLM systems and evaluation tools.
  • Knowledge of AI governance, compliance, and responsible AI practices.
  • Understanding of student psychology and age-appropriate AI output design.

More Info

Job Type:
Industry:
Employment Type:

Key Skills

AI risk assessments

NIST AI RMF

vulnerability identification

AI evaluation frameworks

hallucination grounding

AI governance documentation

child-safe content standards

AI release pipelines

response-quality assessments

responsible AI

India EU AI regulations

adversarial testing

red-teaming

audit trails

safety reports

benchmark testing

About Company