How this analysis was created: This summary and feature list are AI-generated from collected project material and can contain mistakes. Stars, license and language are imported from GitHub. Inclusion does not mean that we have tested or audited this project. Check the source documentation for any feature you depend on. Learn more on our About page.
An LLM-free Multi-dimensional Benchmark for Multi-modal Hallucination Evaluation
The main features of junyangwang0410/amber are: Evaluation Benchmarks, Hallucination Mitigation.
Projects with overlapping indexed features include: masaiahhan/correlationqa — The official repository of the paper "The Instinctive Bias: Spurious Images lead to Hallucination in MLLMs". openkg-org/easydetect — An Easy-to-use Hallucination Detection Framework for LLMs. bcdnlp/faithscore — FaithScore: Fine-grained Evaluations of Hallucinations in Large Vision-Language Models. junyangwang0410/haelm — An automatic MLLM hallucination detection framework. pku-yuangroup/chatlaw — ChatLaw is a specialized large language model legal assistant designed to provide automated consulting and question… datawhalechina/prompt-engineering-for-developers — This project is a technical curriculum and development guide focused on large language model prompt engineering,…
The official repository of the paper "The Instinctive Bias: Spurious Images lead to Hallucination in MLLMs"
FaithScore: Fine-grained Evaluations of Hallucinations in Large Vision-Language Models
An Easy-to-use Hallucination Detection Framework for LLMs.