Loading...
正在加载...
请稍候

T²Agent: A Tool-augmented Multimodal Misinformation Detection Agent with Monte Carlo Tree Search

2026-08-06 15:10

🔍 耿同学打假报告

论文信息

  • 论文来源:T2agent.pdf (arXiv:2505.19768v2 [cs.CL] 17 Nov 2025)
  • 标题:T²Agent: A Tool-augmented Multimodal Misinformation Detection Agent with Monte Carlo Tree Search
  • 作者:Xing Cui, Yueying Zou, Zekun Li, Peipei Li, Xinyuan Xu, Xuannan Liu, Huaibo Huang
  • 期刊/会议:AAAI 2026 (Copyright © 2026, Association for the Advancement of Artificial Intelligence)
  • 发表年份:2026 (预印本提交于 2025 年 11 月)

综合评定:🔴 实锤

详细发现

发现 1:核心实验数据严重自相矛盾(编造的百分比与计算结果对不上)

  • 位置:正文 AMG 实验部分 & Table 2 & Table 5
  • 描述:作者在正文中声称其模型在 AMG 数据集上取得了巨大的提升:“our method achieves improvements of 38.6%, 38.6%, and 39.7% over the MMD-agent when using GPT-4.1-nano, GPT-4o-mini, and GPT-4o, respectively.” 但是,查阅论文后续的 Table 5,T²Agent 的 F1 Score 分别为 0.402, 0.499, 0.510。根据 Table 2 中 MMD-agent 的 F1 基线(0.192, 0.227, 0.306),其实际提升幅度分别为 109%、119%、66%!文中所谓的“38.6%”和“39.7%”完全是凭空捏造的数字。
  • 证据:基础算术验证。0.192 × 1.386 = 0.266 ≠ 0.402。真实提升幅度翻了一倍多,作者文本中报告的提升率完全是瞎编的。
  • 严重程度:🔴

发现 2:表格数据大面积缺失(Excel没关就截图发论文了?)

  • 位置:Table 2 (Comparison with MMD-agent on AMG)
  • 描述:Table 2 旨在展示 T²Agent 与基线模型 MMD-agent 在 AMG 数据集上的对比。然而,在列出了 MMD-agent 的数据(如 0.290, 0.192 等)后,T²Agent(Ours)这一行的所有准确率和 F1 数据竟然全是空白!作者用一个空表格向读者宣告“我们赢了”,但连数据都懒得填上去。
  • 证据:直接提取的 Table 2 内容显示,GPT-4.1-nano MMD-agent 0.290 0.192 下一行的 Ours 后面没有任何数字。
  • 严重程度:🔴

发现 3:时间线与引用严重崩塌(把2024年的模型穿越回2022年)

  • 位置:正文引用 (OpenAI 2022) 与 Reference 列表
  • 描述:论文大量使用了 GPT-4o 作为主干模型(GPT-4o 实际发布于 2024 年 5 月),但正文中引用该模型时,标注的参考文献居然是 (OpenAI 2022)。查阅文末参考文献,赫然写着“OpenAI. 2022