Open LLM eval framework, pytest-style — 50+ research-backed metrics (faithfulness, hallucination, relevance, bias) for unit-testing RAG, chatbots, and agents in CI/CD.
Category: Model Inference & LLM Tooling • Stars: 17K
IT News, How To's & updates
Open LLM eval framework, pytest-style — 50+ research-backed metrics (faithfulness, hallucination, relevance, bias) for unit-testing RAG, chatbots, and agents in CI/CD.
Category: Model Inference & LLM Tooling • Stars: 17K