6 projects
jevals
Agent evals and guardrails that run in one request, built on System One models (Jev, Kev, Laya).
jeval
Evaluate JSON values against expected checks.
llmetrics
A metrics and evaluation library for LLMs
langtests
An evaluation library for NLP models
langsafe
An evaluation library for NLP models
langeval
An evaluation library for NLP models