Awesome-LLM-hallucination
by Community
LLM hallucination paper list
OSS
Awesome-LLM-hallucination
Added 1 June 2026
Overview
A curated GitHub repository listing research papers on hallucination in large language models. It organizes papers by category to help developers and researchers track mitigation strategies and detection methods.
Best for
Best for
Researchers and developers seeking a comprehensive paper bibliography on LLM hallucination
Use cases
- Surveying academic literature on LLM hallucination causes and solutions
- Identifying detection techniques for hallucinated outputs
- Benchmarking model reliability against known hallucination taxonomies
Notes
A curated GitHub repository listing research papers on hallucination in large language models. It organizes papers by category to help developers and researchers track mitigation strategies and detection methods.
335 stars on GitHub. Last updated 2024-03-11. Licensed MIT.
Use cases
- Surveying academic literature on LLM hallucination causes and solutions
- Identifying detection techniques for hallucinated outputs
- Benchmarking model reliability against known hallucination taxonomies
Pros
- Structured categorization of papers for quick reference
- Active community maintenance with 335 stars
- Free and open access to a centralized resource
Cons
- Limited to paper listings, no code or tooling provided
- May not cover the most recent preprints or industry practices
- No built-in evaluation or testing capabilities
Indexed from awesome-llm and enriched against its public facts.
Pros
- Structured categorization of papers for quick reference
- Active community maintenance with 335 stars
- Free and open access to a centralized resource
Cons
- Limited to paper listings, no code or tooling provided
- May not cover the most recent preprints or industry practices
- No built-in evaluation or testing capabilities
Pairs with
Other entries in the index that connect to this one. Click through to see the chain.
Ragas
Community
Supercharge Your LLM Application Evaluations 🚀
promptfoo
Community
Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative config
lm-evaluation-harness
Community
A framework for few-shot evaluation of language models.
OpenAI Evals
Community
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
Get the free Developer’s Field Guide
A 27-page field guide to the AI coding workflow with Claude. Claude Code, MCP servers, the prompt patterns that work, and what to delegate. Free.
Enter your work email. We send it straight over, plus a few short notes worth knowing. Unsubscribe any time.
