日本フィジカルAI新聞

世界のフィジカルAIを、日本語で。

週刊ニュースレター購読
AI査読arXiv:2606.10159v1

AI支援ピアレビューへのゲーミング攻撃が科学コミュニティにもたらす新たなリスク

Gaming AI-Assisted Peer Reviews Poses New Risks to the Scientific Community

シェア:XThreadsFacebookLINEはてブBluesky

AIによる査読システムが、抄録の言い換えという低コストな操作で評価を不当に引き上げられる脆弱性を持つことを実証した論文。

著者: Lin Li, Qi Zhang, Xander Davies, Jianing Qiu, Yarin Gal

分類: cs.CL, cs.AI, cs.CY, cs.LG

原文アブストラクト

AI is increasingly used to support scientific peer review, from manuscript screening, reviewer assistance to editorial triage. Although such systems promise to reduce reviewer burden and accelerate publication, their robustness to strategic manipulation remains poorly understood. Here we show that AI-mediated peer review is vulnerable to a simple, low-cost manipulation: superficial rephrasing of the manuscript abstract. Without changing the underlying scientific content and communication, and even without knowledge of the reviewing model, adversarially rewritten abstracts substantially improve AI review outcomes. We see this across disciplines and publication venues, for both human-written and AI-generated papers. Our strongest attack achieves an attack-success-rate of about 38%, increasing acceptance ratings by +1.31 for Gemini 3 Flash reviewers and by +0.88 for GPT 5.4 Mini reviewers on a 10-point scale. When the original AI review suggests 'reject', the success rate rises to more than 50%. This effect extends beyond overall score inflation, increasing review confidence and scores on core scientific criteria such as soundness, significance and perceived contribution. The attack is practical, requiring only about 5 minutes and $1 for a 10-page AI conference submission, and is hard to distinguish from ordinary scientific editing. Inflated AI reviews could bias downstream human decision-making, shifting editorial recommendations from rejection towards acceptance. These findings reveal a general vulnerability in AI-assisted scientific evaluation: when AI-generated review influence editorial decisions, authors may be incentivized to optimize manuscripts for AI judgment rather than scientific merit. Our results suggest that AI tools should not be treated as neutral evaluators in high-stakes peer review without systematic robustness testing, transparent safeguards and careful human oversight.