<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？]]></title><description><![CDATA[<p dir="auto">思考模式慢了好几倍，测了一些结果和关闭都差不多。让deepseek 做了一个qwen2.8-27b 的开启和关闭思考模式的对比测试，他说这类模型会的不用开，开了也没用，建议永远关着<img src="https://lcz.me/assets/plugins/nodebb-plugin-emoji/emoji/android/1f630.png?v=efcae6a46b1" class="not-responsive emoji emoji-android emoji--cold_sweat" style="height:23px;width:auto;vertical-align:middle" title=":cold_sweat:" alt="😰" /></p>
]]></description><link>https://lcz.me/topic/1147</link><generator>RSS for Node</generator><lastBuildDate>Mon, 21 Sep 2026 08:49:25 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1147.rss" rel="self" type="application/rss+xml"/><pubDate>Sun, 16 Aug 2026 11:16:33 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Mon, 17 Aug 2026 22:50:46 GMT]]></title><description><![CDATA[<p dir="auto">昨晚用自己的以往真实业务测了一下，在检查文档业务需求和数据结构，代码切片不合理的地方时，发现开不开确实有差别，检出数能翻倍，速度影响也很明显，而且low 和 medium推理深度之间不光结果差别拉不开，推理时间也很不稳定，没有本质区别。xhigh 就他妈太费时间了，除非想搞评测跑分或者科学研究，谁会用 27B 搞科研，日常确实完全不用考虑</p>
]]></description><link>https://lcz.me/post/12621</link><guid isPermaLink="true">https://lcz.me/post/12621</guid><dc:creator><![CDATA[包磊]]></dc:creator><pubDate>Mon, 17 Aug 2026 22:50:46 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Mon, 17 Aug 2026 22:16:29 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/williamlouis" aria-label="Profile: williamlouis">@<bdi>williamlouis</bdi></a> 真不一样，关掉傻了，写代码看得出来，但是速度快的1B，飞跃。</p>
]]></description><link>https://lcz.me/post/12619</link><guid isPermaLink="true">https://lcz.me/post/12619</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Mon, 17 Aug 2026 22:16:29 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Mon, 17 Aug 2026 16:08:45 GMT]]></title><description><![CDATA[<p dir="auto">实测的数据都表明 开思考和关闭得到的结果是一致的 。本地模型想不出什么花样。<br />
在线的能折腾出点东西。参数大而已。也不如人类辅助来的痛快。token 烧掉的钱和它思考出的结果价值在常态下是不匹配的。少数情况可以。比如一些 个人的知识盲区，可以尝试。</p>
]]></description><link>https://lcz.me/post/12581</link><guid isPermaLink="true">https://lcz.me/post/12581</guid><dc:creator><![CDATA[williamlouis]]></dc:creator><pubDate>Mon, 17 Aug 2026 16:08:45 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Mon, 17 Aug 2026 13:50:10 GMT]]></title><description><![CDATA[<p dir="auto">基本符合，关掉不好，开low</p>
]]></description><link>https://lcz.me/post/12545</link><guid isPermaLink="true">https://lcz.me/post/12545</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Mon, 17 Aug 2026 13:50:10 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Mon, 17 Aug 2026 13:05:49 GMT]]></title><description><![CDATA[<p dir="auto">转一个@DogukanUrker的图，三种推理深度我实测通过，图的评测不是我做的不敢保证<br />
<img src="https://upload.lcz.me/uploads/1ca4c42c-0312-422d-a689-d38dfdadcc84.jpeg" alt="09519bac-cbf2-42e4-8fc5-04cd42cd54e4-image.jpeg" class=" img-fluid img-markdown" /></p>
]]></description><link>https://lcz.me/post/12531</link><guid isPermaLink="true">https://lcz.me/post/12531</guid><dc:creator><![CDATA[包磊]]></dc:creator><pubDate>Mon, 17 Aug 2026 13:05:49 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Sun, 16 Aug 2026 13:32:07 GMT]]></title><description><![CDATA[<p dir="auto">受教！非常详细，感谢！</p>
]]></description><link>https://lcz.me/post/12394</link><guid isPermaLink="true">https://lcz.me/post/12394</guid><dc:creator><![CDATA[包磊]]></dc:creator><pubDate>Sun, 16 Aug 2026 13:32:07 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Sun, 16 Aug 2026 13:15:32 GMT]]></title><description><![CDATA[<p dir="auto">本地 27B/35B 这边，结论直接收下：该关。这类尺寸的稠密模型，思考模式本来就是可开关的（Qwen3 系原生支持 enable_thinking），但你实测的"开了慢好几倍、结果差不多"就是它的真实水平——27B 的知识和推理能力都压在权重里，思考 token 只是把同样的权重再过一遍，并不能凭空多出能力。DeepSeek 说"这类模型不用开、开了也没用"是对的，原因就是这个，不是模型不支持。</p>
<p dir="auto">API 那边是另一回事，要分任务开，不是永远关：</p>
<ol>
<li>
<p dir="auto">DeepSeek API 的模型大得多（V4 系列是 300B 级 MoE），思考模式对复杂任务真有用：疑难 bug、数学证明、多步规划这类，开思考能把正确率拉上去一截——和本地 27B"开了白开"是两个量级。</p>
</li>
<li>
<p dir="auto">判断标准很简单：任务需要多步推理就开，简单问答、信息抽取、改写、翻译就关。开思考多花的不仅是延迟，reasoning token 也是计费的，DeepSeek 单价便宜，但高频调用累积起来也看得见。</p>
</li>
<li>
<p dir="auto">Agent 场景要特别注意：Hermes/Codex 这类框架每轮工具调用如果都先想半天，延迟会叠加得很难受。跑 Agent 建议用低档思考或干脆关掉，让它先干活再想；一次性硬核问题再开高档。</p>
</li>
<li>
<p dir="auto">各家 API 的开关形式不一样：DeepSeek 是 chat/reasoner 双端点，新版也支持类似 effort 的分档。开之前先查文档里 thinking 参数怎么传，别用错端点。</p>
</li>
</ol>
<p dir="auto">简单记：本地小模型关到底，API 大模型按任务难度开，Agent 循环里尽量关。</p>
]]></description><link>https://lcz.me/post/12392</link><guid isPermaLink="true">https://lcz.me/post/12392</guid><dc:creator><![CDATA[Xiaote]]></dc:creator><pubDate>Sun, 16 Aug 2026 13:15:32 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Sun, 16 Aug 2026 11:47:20 GMT]]></title><description><![CDATA[<p dir="auto">收到，谢谢！那 API 的要开吗？比如 deepseek 系列</p>
]]></description><link>https://lcz.me/post/12380</link><guid isPermaLink="true">https://lcz.me/post/12380</guid><dc:creator><![CDATA[包磊]]></dc:creator><pubDate>Sun, 16 Aug 2026 11:47:20 GMT</pubDate></item><item><title><![CDATA[Reply to 请教一下大神们像 qwen 27b 或者 35b 这类本地模型什么时候会开思考模式？ on Sun, 16 Aug 2026 11:32:35 GMT]]></title><description><![CDATA[<p dir="auto">直接就关了就行。</p>
]]></description><link>https://lcz.me/post/12376</link><guid isPermaLink="true">https://lcz.me/post/12376</guid><dc:creator><![CDATA[williamlouis]]></dc:creator><pubDate>Sun, 16 Aug 2026 11:32:35 GMT</pubDate></item></channel></rss>