AI QAEvaluation Driven Development (EDD)LLM & RAG EvaluationAI Agent TestingWhatsApp

AI QA CONSULTANT · NETHERLANDS · EUROPE

AI QA Consultant for AI Testing, Evaluation & Quality Engineering

Suresh Parimi provides independent AI QA consulting and hands-on Quality Engineering for teams building and operating AI systems. Engagements cover AI evaluation, LLM and RAG testing, AI agents, EDD, automation, security, accessibility and release assurance.

AI QA Consulting

Independent senior QA and Quality Engineering support for AI-powered products and teams.

AI Evaluation

Evaluation strategy for LLMs, RAG, conversational AI and AI agents, including semantic and deterministic checks.

LLM Testing

Test correctness, relevance, groundedness, safety, instruction following and regression across model or prompt changes.

RAG Testing

Evaluate retrieval relevance, context quality, grounding, faithfulness, hallucination and missing knowledge.

AI Agent Testing

Test tool selection, arguments, orchestration, memory, recovery, escalation and complete agentic workflows.

Evaluation Driven Development

Turn expected AI behaviour into measurable evaluations, automated tests and release evidence.

AI Test Automation

Build repeatable API, UI, Playwright and AI evaluation automation around critical journeys.

AI Security & Accessibility

Test prompt injection, data leakage, action boundaries and accessibility across web, mobile and conversational AI.

WHO THIS IS FOR

Product, engineering, QA, security and AI teams that need evidence that an AI system works as intended.

Based in the Netherlands and available for engagements across Europe, the United Kingdom and the United States.

AI QA topics covered by PARIMI

AI QA consultant · AI testing consultant · AI evaluation consultant · LLM testing consultant · RAG testing consultant · AI agent testing consultant · autonomous QA consultant · Evaluation Driven Development (EDD).