ENTERPRISE AI QUALITY ENGINEERING

GenAI Test Lead

LLM Testing • RAG Validation • AI Automation • Responsible AI

🤖 Enterprise AI • 8+ Years Experience • Full-Time

LLM Testing RAG AI Automation Test Leadership

About the Role

We are looking for an experienced GenAI Test Lead to lead end-to-end testing and quality assurance for enterprise Generative AI applications. This role combines AI quality engineering, automation, customer engagement, governance, and production AI validation.

💼 Experience
8+ Years
🤖 Domain
Generative AI Testing
👨‍💻 Role
AI Test Lead
🏢 Employment
Full-Time

Role Overview

Lead enterprise AI testing initiatives across Large Language Models, AI agents, conversational AI, and Retrieval-Augmented Generation (RAG) systems. Drive AI quality strategy from planning through production release while ensuring compliance with responsible AI standards.

Key Responsibilities

  • Lead end-to-end AI testing strategy, planning, execution, and release management.
  • Define testing strategies for LLMs, RAG systems, conversational AI, and AI agents.
  • Perform AI model evaluation including hallucination, bias, fairness, robustness, and explainability testing.
  • Manage sprint planning, QA delivery, quality metrics, and release readiness.
  • Serve as the primary customer contact for AI testing activities.
  • Create AI quality dashboards and executive-level reports.
  • Implement AI governance aligned with ISO/IEC 42001, NIST AI RMF, and EU AI Act.
  • Mentor QA engineers on AI testing, automation, and Python best practices.

Required Skills

LLM Testing Prompt Engineering RAG AI Agents DeepEval Ragas LangSmith Promptfoo Python PyTest Playwright Selenium API Testing GitHub Actions Jenkins AWS Azure GCP

Must Have

  • 8+ years in Software Testing and Quality Assurance.
  • 1–2 years of hands-on AI/ML and GenAI testing experience.
  • Experience leading QA teams and customer-facing testing engagements.
  • Strong Python automation skills using PyTest.
  • Experience with Selenium, Playwright, API testing, and Postman.
  • Hands-on LLM evaluation and prompt testing experience.
  • Knowledge of DeepEval, Ragas, LangSmith, or Promptfoo.
  • Understanding of AI governance and Responsible AI standards.

Technology Stack

  • Python
  • PyTest
  • Playwright
  • Selenium
  • Requests
  • Postman
  • DeepEval
  • Ragas
  • LangSmith
  • Promptfoo
  • GitHub Actions
  • Jenkins
  • AWS
  • Azure
  • Google Cloud
  • Jira
  • Zephyr
  • TestRail

AI Testing Standards

  • ISO/IEC 42001
  • NIST AI Risk Management Framework (AI RMF)
  • EU AI Act
  • Responsible AI
  • ISTQB CT-AI

Why Join?

Take ownership of enterprise AI quality strategy while working alongside AI architects, prompt engineers, DevOps specialists, and machine learning experts. Lead testing programs that directly impact production AI systems used by customers worldwide.

Highlights

  • ✔ End-to-End AI Quality Leadership
  • ✔ Enterprise LLM Testing
  • ✔ RAG & AI Agent Validation
  • ✔ Responsible AI Governance
  • ✔ Python Automation Frameworks
  • ✔ Customer-Facing Leadership Role
  • ✔ Continuous Learning in AI Engineering

Disclaimer: This job information is shared for educational and career guidance purposes only. Company names and trademarks belong to their respective owners. Please apply through the official company careers page or LinkedIn job posting.