VERIFYAI infra
Typed Evals: LLM Evaluation Framework
Typed Evals is an open-source Python framework designed for evaluating LLMs, RAG pipelines, and AI agents by leveraging typed judge backends such as Jev.
Jev-as-a-judge for LLM/Agents
⚡ Typed Evals — an open-source Python framework for evaluating LLMs, RAG pipelines, and AI agents using System One Models like Jev and other typed judge backends.
GitHub: github.com/TrustifAI/typed_evals
