MIT-licensed Framework for LLMs, RAGs, Chatbots testing. Configurable via YAML and integrable into CI pipelines for automated testing.
-
Updated
Dec 11, 2024 - Python
MIT-licensed Framework for LLMs, RAGs, Chatbots testing. Configurable via YAML and integrable into CI pipelines for automated testing.
Library for Microsoft Bot Framework Chatbot unit testing
Sub-second RAG regression testing. Define golden questions, detect lost chunks in CI. pytest for your RAG pipeline.
EvalBot — local-first chatbot security & quality evaluation (FastAPI + Next.js). Evaluate chatbot answers against your own docs & guidelines with ML/NLP + AI-judge scoring. Apache-2.0.
An automated approach for exploring and testing conversational agents using large language models. TRACER discovers chatbot functionalities, generates user profiles, and creates comprehensive test suites for conversational AI systems.
pytest lab for testing LLMs: RAG eval, red teaming, guardrails, drift monitoring — 14 modules, 382 tests, zero API calls needed
A Python library to connect and interact with chatbots.
QA framework for testing conversational AI systems (LLM agents, chatbots, voice assistants) with workflow validation and regression checks
An open-source framework for robust, LLM-powered testing and tracing of conversational AI applications.
A plug & play framework for generative ai projects to be tested & automated
A framework for testing LLM-based chatbots in regulated industries (telco, banking, insurance). Covers hallucination detection, prompt injection resistance, response quality scoring and regression testing.
UI for persona-api
Bilingual portfolio project for evaluating chatbot helpfulness, accuracy, tone, safety, and instruction following.
A Node.js testing framework for ChatBots
Deterministic hosted and customer-executed local testing for AI agents
Modular, extensible QA framework for evaluating AI chatbots — built for CI/CD pipelines, model comparison, and continuous quality monitoring.
Quality auditor for AI chatbots. Analyzes your conversation logs to show where the bot is underperforming.
AI Chatbot Testing Project showcasing QA documentation, test scenarios, test cases, defect reporting, test execution and testing practices.
QA & test automation case study - the shared AI-powered dispute resolution layer behind six connected fintech products, resolving ~80% of disputes without a human agent. Playwright + TypeScript automation, AI/chatbot testing.
Evaluation study of a Spanish telecoms or energy support chatbot: 61 bilingual test cases, a severity-weighted rubric and a model judge checked against human scoring
To associate your repository with the chatbot-testing topic, visit your repo's landing page and select "manage topics."