llm-eval
A flexible, extensible, and reproducible framework for evaluating LLM workflows, applications, retrieval-augmented generation pipelines, and standalone models across custom and standard datasets.
// repository documentation
Was this content helpful?
(0 ratings)
