PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 15, 20242 citationsOpen Access

Mitigating Dialogue Hallucination for Large Multi-modal Models via Adversarial Instruction Tuning

View Full Paper
DPDongmin ParkZQZhaofang QianGHGuangxing Han

Key Points

Key points are not available for this paper at this time.

Abstract

Mitigating hallucinations of Large Multi-modal Models(LMMs) is crucial to enhance their reliability for general-purpose assistants. This paper shows that such hallucinations of LMMs can be significantly exacerbated by preceding user-system dialogues. To precisely measure this, we first present an evaluation benchmark by extending popular multi-modal benchmark datasets with prepended hallucinatory dialogues generated by our novel Adversarial Question Generator, which can automatically generate image-related yet adversarial dialogues by adopting adversarial attacks on LMMs. On our benchmark, the zero-shot performance of state-of-the-art LMMs dropped significantly for both the VQA and Captioning tasks. Next, we further reveal this hallucination is mainly due to the prediction bias toward preceding dialogues rather than visual content. To reduce this bias, we propose Adversarial Instruction Tuning that robustly fine-tunes LMMs on augmented multi-modal instruction-following datasets with hallucinatory dialogues. Extensive experiments show that our proposed approach successfully reduces dialogue hallucination while maintaining or even improving performance.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Park et al. (2024) studied this question.

synapsesocial.com/papers/68e73dcfb6db6435876b74e0https://doi.org/10.48550/arxiv.2403.10492
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization2025
  2. 2Mitigating Object Hallucination via Data Augmented Contrastive Tuning2024 · 1 citations
  3. 3Prescribing the Right Remedy: Mitigating Hallucinations in Large Vision-Language Models via Targeted Instruction Tuning2024
  4. 4Constructing Benchmarks and Interventions for Combating Hallucinations in LLMs2024
  5. 5Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models2024 · 1 citations