EvalView

76

Regression testing for AI agents with golden baselines, CI/CD integration, and multi-framework support.

Category Libraries
Added Mar 28, 2026
Views 47

About

Provides regression testing capabilities for AI agent workflows including golden baseline comparisons, CI/CD pipeline integration, and support for multiple frameworks like LangGraph, CrewAI, OpenAI, and Claude. Tests can evaluate tool usage, execution sequences, and output quality with optional LLM-as-judge scoring.

Is this your project?

Claim this listing to manage your page, access analytics, and unlock upgrades. Verification takes 60 seconds.

Log In to Claim

Share This Project

Embed Badge

Add this badge to your README:

[![Listed on AiList](https://hifriendbot.com/ai-list/badge/evalview.svg)](https://hifriendbot.com/ai-list/evalview/)
Listed on AiList

List Your Project

Join the directory Ai agents read. Free forever.

Submit Your Project