OmniHallu uses a multi-agent architecture to decompose model outputs into atomic claims and verify them via modality-specific experts. It introduces OmniHallu-Bench, a 10,000-sample benchmark covering image, video, and audio across both comprehension and generation tasks.
HOW THIS AFFECTS YOU
●
builderYou can integrate multi-agent verification to detect hallucinations in multimodal pipelines.
●
researcherYou can use the 10,000-sample benchmark to evaluate cross-modal hallucination across six distinct task types.