Multimodal Learning 相关度: 9/10

SciFigDetect: A Benchmark for AI-Generated Scientific Figure Detection

You Hu, Chenzhuo Zhao, Changfa Mo, Haotian Liu, Xiaobai Li
arXiv: 2604.08211v1 发布: 2026-04-09 更新: 2026-04-09

AI 摘要

提出了SciFigDetect,一个用于检测AI生成科学图像的新基准,揭示了现有检测方法的不足。

主要贡献

  • 构建了首个AI生成科学图像检测基准SciFigDetect
  • 设计了基于代理的数据生成流程,涵盖多种图像类型和生成源
  • 评估了现有检测器在零样本、跨生成器和图像退化情况下的表现,发现其存在显著缺陷

方法论

使用基于代理的数据管道检索论文、理解文本和图像,构建prompt生成候选图像,并进行过滤和优化。

原文摘要

Modern multimodal generators can now produce scientific figures at near-publishable quality, creating a new challenge for visual forensics and research integrity. Unlike conventional AI-generated natural images, scientific figures are structured, text-dense, and tightly aligned with scholarly semantics, making them a distinct and difficult detection target. However, existing AI-generated image detection benchmarks and methods are almost entirely developed for open-domain imagery, leaving this setting largely unexplored. We present the first benchmark for AI-generated scientific figure detection. To construct it, we develop an agent-based data pipeline that retrieves licensed source papers, performs multimodal understanding of paper text and figures, builds structured prompts, synthesizes candidate figures, and filters them through a review-driven refinement loop. The resulting benchmark covers multiple figure categories, multiple generation sources and aligned real--synthetic pairs. We benchmark representative detectors under zero-shot, cross-generator, and degraded-image settings. Results show that current methods fail dramatically in zero-shot transfer, exhibit strong generator-specific overfitting, and remain fragile under common post-processing corruptions. These findings reveal a substantial gap between existing AIGI detection capabilities and the emerging distribution of high-quality scientific figures. We hope this benchmark can serve as a foundation for future research on robust and generalizable scientific-figure forensics. The dataset is available at https://github.com/Joyce-yoyo/SciFigDetect.

标签

AI生成图像检测 科学图像 基准数据集 视觉取证 多模态学习

arXiv 分类

cs.CV