github Actively maintained

vectara/hallucination-leaderboard

Leaderboard Comparing LLM Performance at Producing Hallucinations when Summarizing Short Documents

2 awesome lists

Quick read

Stars
3,312
Forks
107
Open issues
19
Commits
452

Activity and growth

Latest capture 2026-09-10 10:55

Stars · last 7 days
0 0.0%
Commits · last 7 days
0 0.0%
Stars since tracking
0
Stored snapshots
1

Classification

Metadata

Language
Python
License
Apache-2.0
Default branch
main
Created
2023-10-31
First commit
2023-10-31
Last pushed
2026-05-11
GitHub updated
2026-09-10
Last synced
2026-09-10 10:55
Stack scanned
2026-09-10 10:55
Archived
No

AI development signals

0 paths

Agent instructions and tool configuration found in this repository.

No config files detected.

Growth history

Tracked growth

1 observed capture since 2026-09-10. Observed captures are shown by default.

Stars from first capture 0

Chart data

Observed captures only

Time horizon

All tracked data

Custom date range

Stars history

Observed snapshots

Commits history

Observed snapshots

Similar repositories

Nearest indexed repositories by embedding similarity.

FranxYao/chain-of-thought-hub

Benchmarking large language models' complex reasoning ability with chain-of-thought prompting

2,774 stars
Jupyter Notebook 2 awesome lists

EvolvingLMMs-Lab/lmms-eval

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

4,396 stars
Python 1 awesome list

GPT-Fathom/GPT-Fathom

GPT-Fathom is an open-source and reproducible LLM evaluation suite, benchmarking 10+ leading open-source and closed-source LLMs as well as OpenAI's earlier models on 20+ curated benchmarks under aligned settings.

344 stars
Python 1 awesome list

evalplus/evalplus

Rigourous evaluation of LLM-synthesized code - NeurIPS 2023 & COLM 2024

1,803 stars
Python 2 awesome lists