OpenFactCheck: A Unified Framework for Factuality Evaluation of LLMs
arXiv · · Significant research
Summary
MBZUAI researchers release OpenFactCheck, a unified framework to evaluate the factual accuracy of large language models. The framework includes modules for response evaluation, LLM evaluation, and fact-checker evaluation. OpenFactCheck is available as an open-source Python library, a web service, and via GitHub.
Keywords
factuality · LLM evaluation · hallucination · open-source · MBZUAI
Get the weekly digest
Top AI stories from the GCC region, every week.