aifollow.news 搜索
返回 Rohan Paul
Rohan Paul· @rohanpaul_ai · X· · 原发布时间 AI 评分60

研究发现:新一代模型可能让 AI 文本检测器漏检论文摘要

自动核验发布 · 本文由系统生成并完成证据核验,未经人工审稿。

AI 导读

东京首都大学一项研究发现,AI 文本检测器的识别效果随改写模型版本而变化:一组检测器在模型换代前能识别超过 99% 的改写文本,换代后仅识别 3.8%。另一项结果显示,Pangram 漏检了 Meta Muse-Glimmer 改写的 79.8% 的科学摘要。依赖这类工具筛查论文的机构需在模型更新后重新测试。

正文 · 原文

该语言的正文暂不可用,当前显示已有版本。

@rohanpaul_ai

This paper finds that AI-text detectors trained on a vendor's older models caught over 99% of rewrites before a generation change and only 3.8% after it.

Detectors that screen scientific papers for AI writing can stop working when a new LLM generation arrives, so anyone relying on them should re-test them with every model release.

引用Rohan Paul@rohanpaul_ai
Pangram, the AI-text detector, missed 79.8% of scientific abstracts rewritten by Meta's Muse-Glimmer, while flagging just 1 of 5,000 human abstracts. In a new paper from Tokyo Metropolitan University, reseaerchers find the share of AI-rewritten abstracts that Pangram misses depends strongly on the LLM version Shows that it caught 93.5% of GPT-5 rewrites but missed 79.8% from another new model. Its miss rate depended mostly on which model did the rewriting.
在 X 查看这条帖子 ↗

来源:Rohan Paul · x.com

论文
发现内容有误?提交纠错