DeepEval:A comprehensive benchmark for evaluating Large Multimodal Models' capacities of visual deep semantics.
py-deepeval-behave-bdd-testing-example:An example that combines Behave (BDD testing) with DeepEval (LLM evaluation) to create human-readable, stakeholder-friendly tests for AI Agents / chatbots.
// repository documentation
Was this content helpful?
★ 0(0 ratings)
Recent Feedback
Download README
Do you want to download the README.md file for deepeval-wrapper?