跳到正文
Rohan Paul· @rohanpaul_ai · X·· 2 小时前AI 评分60
AI 导读

清华大学、牛津大学与斯坦福大学的新论文发现,LLM 即使关闭思考仍会继续输出推理过程,在开放式问题上尤为明显。关闭思考后,DeepSeek-V4-Flash 在 99.9% 的开放式回答中仍写出推理内容,缺少 think 标签或回复简短都不能证明模型跳过了推理。

正文

New Tsinghua, Oxford, and Stanford paper finds that LLMs keep reasoning out loud even with thinking turned off, especially on open-ended questions.

With thinking disabled, DeepSeek-V4-Flash still wrote out reasoning in 99.9% of open-ended answers. A missing think tag or a short reply does not prove the model skipped reasoning.

Yes/no questions are easy to answer directly, multiple-choice questions sit in the middle, and open-ended ones keep pulling the model back into reasoning.

Forcing answer-only replies on open-ended tasks raised compliance to about 40% across 5 models but cut accuracy by about 15 points.

– arxiv. org/abs/2610.11765

Title: "Thinking Inertia: LLMs Keep Thinking When Told Not To"

来源:Rohan Paul · x.com