EvalView
76
Regression testing for AI agents with golden baselines, CI/CD integration, and multi-framework support.
About
Provides regression testing capabilities for AI agent workflows including golden baseline comparisons, CI/CD pipeline integration, and support for multiple frameworks like LangGraph, CrewAI, OpenAI, and Claude. Tests can evaluate tool usage, execution sequences, and output quality with optional LLM-as-judge scoring.
Is this your project?
Claim this listing to manage your page, access analytics, and unlock upgrades. Verification takes 60 seconds.
Share This Project
Embed Badge
Add this badge to your README:
[](https://hifriendbot.com/ai-list/evalview/)
