The Eval Index / RAG Eval / #168
multivon-ai/multivon-eval
by multivon-ai · RAG Eval · updated 23d ago
Practical LLM evaluation for teams that ship to production. Deterministic + LLM-as-judge evaluators, dataset support, CI/CD integration.
45
momentum
25
stars
0
forks
#168
rank
agent-evaluationai-evaluationevalshallucination-detectionllm-as-judgellm-evalllm-evaluationllmopsmlopsprompt-engineeringpythonrag-evaluation
View on GitHub →