Open-source workspace where domain experts and engineers build test sets for AI agents together.
Rhesis AI is a testing platform for LLM and agent applications that brings non-technical domain experts into the same workspace as engineers, so both can write test scenarios instead of leaving coverage entirely to developers. It generates test cases including adversarial and edge-case prompts, runs them against an application, and traces failures back to a root cause. A red-teaming component called Polyphemus probes for jailbreaks, prompt injection, and PII leakage. The core platform is open source under MIT and self-hostable; an Enterprise tier adds SSO, role-based access, and API clients, though its price is not published.
The Community edition is free and open source (MIT); Enterprise pricing (SSO, RBAC, API clients) is not published.
Use tool ↗Rhesis is worth a look for teams that want domain experts, not just engineers, writing test cases for an AI agent, and who are comfortable self-hosting an open-source tool. The MIT-licensed core is genuinely free, but anything past the community tier requires talking to sales since Enterprise pricing isn't posted.
No reviews yet — be the first to review Rhesis AI.
Reviews are tied to your EffectHub account — one review per tool, so the rating for Rhesis AI reflects real users.
No questions yet — ask the first one about Rhesis AI.
Questions about Rhesis AI are tied to your EffectHub account, so answers can reach you and the section stays free of spam.