Rhesis AI
Open-source testing platform for LLM and agentic applications. Test generation, adversarial probing, and regression tracking.
Rhesis AI is an open-source (MIT) testing platform for LLM and AI agent applications. It goes beyond scoring to provide a full testing workflow: define requirements, generate test scenarios (including edge cases), execute them, review results, and track fixes. Engineers work in the SDK, non-technical team members work in the UI.
The platform generates test scenarios from requirements and can connect to knowledge sources via MCP (Notion, GitHub, Jira). Its red-teaming agent, Polyphemus, continuously probes for jailbreaks, prompt injection, and PII extraction. Every failed test links to its root cause across multi-step and multi-agent flows.
Covers conversational AI, RAG, NL-to-SQL, and agentic systems across any LLM provider.
Pricing: Free
Rhesis AI Alternatives
Explore 45 products in the Observability & Analytics category. View all Rhesis AI alternatives.
RAGAS
Open-source evaluation and testing framework for LLM and RAG applications
Guardrails AI
Open-source framework for adding input and output validators around LLM calls
NeMo Guardrails
NVIDIA toolkit for adding programmable guardrails to LLM conversational apps
Presidio
Microsoft open-source SDK for detecting and anonymizing PII in text and images
Work on Rhesis AI? Feature it at the top of Observability & Analytics.
Is your product missing?