aifollow.news 搜索
返回 Rohan Paul
Rohan Paul· @rohanpaul_ai · X· · 原发布时间 AI 评分54

清华、牛津与斯坦福论文发现:关闭思考后,大模型仍常写出推理过程

自动核验发布 · 本文由系统生成并完成证据核验,未经人工审稿。

AI 导读

论文称,关闭思考后,DeepSeek-V4-Flash 在 99.9% 的开放式问题回答中仍写出推理过程;问题越开放,越容易出现这一现象。对5个模型强制要求只给答案,将开放式任务的遵从率提高到约40%,但准确率下降约15个百分点。

正文 · 原文

该语言的正文暂不可用,当前显示已有版本。

New Tsinghua, Oxford, and Stanford paper finds that LLMs keep reasoning out loud even with thinking turned off, especially on open-ended questions.

With thinking disabled, DeepSeek-V4-Flash still wrote out reasoning in 99.9% of open-ended answers. A missing think tag or a short reply does not prove the model skipped reasoning.

Yes/no questions are easy to answer directly, multiple-choice questions sit in the middle, and open-ended ones keep pulling the model back into reasoning.

Forcing answer-only replies on open-ended tasks raised compliance to about 40% across 5 models but cut accuracy by about 15 points.

– arxiv. org/abs/2610.11765

Title: "Thinking Inertia: LLMs Keep Thinking When Told Not To"

来源:Rohan Paul · x.com

论文
发现内容有误?提交纠错