Back to results

Virginia Tech

M3D: Multimodal MultiDocument Fine-Grained Inconsistency Detection

Abstract

dc:description.abstract

Validating claims from misinformation is a highly challenging task that involves understanding how each factual assertion within the claim relates to a set of trusted source materials. Existing approaches often make coarse-grained predictions but fail to identify the specific aspects of the claim that are troublesome and the specific evidence relied upon. In this paper, we introduce a method and new benchmark for this challenging task. Our method predicts the fine-grained logical relationship of each aspect of the claim from a set of multimodal documents, which include text, image(s), video(s), and audio(s). We also introduce a new benchmark (M^3DC) of claims requiring multimodal multidocument reasoning, which we construct using a novel claim synthesis technique. Experiments show that our approach significantly outperforms state-of-the-art baselines on this challenging task on two benchmarks while providing finer-grained predictions, explanations, and evidence.

Degree

thesis:*
Name thesis:degree_name
Master of Science
Level thesis:degree_level
masters
Discipline thesis:degree_discipline
Computer Science & Applications
Department dc:contributor.department
Computer Science and#38; Applications
Grantor dc:publisher
Virginia Tech
Year dc:date.issued
2024

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Tang, Chia-Wei
Chair dc:contributor.committeechair
  • Thomas, Christopher Lee
Committee members dc:contributor.committeemember
  • Lourentzou, Ismini
  • Huang, Lifu

Subjects

dc:subject × 7

Rights

dc:rights
Statement dc:rights
  • In Copyright
Language dc:language.iso
en

Identifiers

dc:identifier.*
Dc Identifier Other
vt_gsexam:40024
OAI identifier oai:identifier
oai:vtechworks.lib.vt.edu:10919/119382

Chain of custody

source
Harvested from
Virginia Tech
Base URL
vtechworks.lib.vt.edu/oai/request
Last updated
2026-07-22
Source record
OAI-PMH GetRecord
citation

Tang, Chia-Wei. M3D: Multimodal MultiDocument Fine-Grained Inconsistency Detection. masters thesis, Virginia Tech, 2024. https://hdl.handle.net/10919/119382