<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[Qwen3.8-27B 关掉推理～效率快太多 成果差不多]]></title><description><![CDATA[<p dir="auto">推理真的没你想像中的有用 至少在我使用的过程中制造不少问题</p>
<p dir="auto">使用的 <a href="https://github.com/syv-ai/qwen38-27b-rtx3090" rel="nofollow ugc">syv-ai/qwen38-27b-rtx3090</a>, 单卡3090长上下文 240k 设置, 不是在 VLLM 设定关闭推理</p>
<p dir="auto">而是在 omp agent harness 上关闭推理 直接设定 off, coding 部分成果差不多 該錯還是會錯，但省下的時間跟上下文真的太多！<br />
我测试的这专案也不单纯，是在原生开源的基础上新增 herdr 的对接 跟 灯光效果设定，在硬体测试上至少错了五次（但这是AI本身就没办法克服的硬件韧体测试，一定要有人类观察结果反餽），所以结论建议：性子急的可以关掉推理了！ 真的快太多！</p>
<p dir="auto">测试专案 repo:<a href="https://github.com/botio/duckyPad-herdr/" rel="nofollow ugc">https://github.com/botio/duckyPad-herdr/</a></p>
]]></description><link>https://lcz.me/topic/1466</link><generator>RSS for Node</generator><lastBuildDate>Wed, 09 Sep 2026 22:53:45 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1466.rss" rel="self" type="application/rss+xml"/><pubDate>Wed, 02 Sep 2026 07:58:41 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 09 Sep 2026 10:48:03 GMT]]></title><description><![CDATA[<p dir="auto">qwen3.8 把thinking 弄成这样乱78糟肯定有他的用处，我关了影响不大，又开回了。。<br />
最近更新了我认为好用的nvfp4无审查版本，只是开动q8 kv &amp; 152000k, 还没到oom 又有速度 （75ts),prefill 2800ts 我就不大爱管他，我使用我自己的自制api, loading model 后就开始跑，不大会thinking，然而使用dsh 就会thinking 来thinking 去老半天。。。还没捣鼓原理在哪</p>
]]></description><link>https://lcz.me/post/16874</link><guid isPermaLink="true">https://lcz.me/post/16874</guid><dc:creator><![CDATA[imbiplaza ASUS]]></dc:creator><pubDate>Wed, 09 Sep 2026 10:48:03 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 09 Sep 2026 08:44:39 GMT]]></title><description><![CDATA[<p dir="auto">使用這個 <a href="https://huggingface.co/froggeric/Qwen-Fixed-Chat-Templates" rel="nofollow ugc">https://huggingface.co/froggeric/Qwen-Fixed-Chat-Templates</a><br />
再設在low 或 medium的效果還不錯。</p>
]]></description><link>https://lcz.me/post/16868</link><guid isPermaLink="true">https://lcz.me/post/16868</guid><dc:creator><![CDATA[paul hou]]></dc:creator><pubDate>Wed, 09 Sep 2026 08:44:39 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 09 Sep 2026 08:38:55 GMT]]></title><description><![CDATA[<p dir="auto">留給機器思考，浪費時間，也不可能做到零錯誤<br />
不如快點生出來快點測，早點發現問題繼續罵他個祖宗十八代<br />
就跟圍棋下快棋一樣，高手落子快還是高手，資質不足給他再多時間依舊廢物</p>
]]></description><link>https://lcz.me/post/16867</link><guid isPermaLink="true">https://lcz.me/post/16867</guid><dc:creator><![CDATA[dardeaw feng]]></dc:creator><pubDate>Wed, 09 Sep 2026 08:38:55 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 14:49:00 GMT]]></title><description><![CDATA[<p dir="auto">看工作场景吧，编程，还有对账，涉及资金之内的，不差时间我都开最大思考</p>
]]></description><link>https://lcz.me/post/15522</link><guid isPermaLink="true">https://lcz.me/post/15522</guid><dc:creator><![CDATA[stxpnet]]></dc:creator><pubDate>Wed, 02 Sep 2026 14:49:00 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 12:30:21 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/terry" aria-label="Profile: terry">@<bdi>terry</bdi></a> 是的，现在已经走 less is more 的路线了， skill 太多 废话太多 都是负担， 现在越聪明的MODEL 越不需要太多额外的提示词，只要给明确的目标，途中不要走歪就好，一开始MAP的路线非常重要，现在个人训练出的 harness 是让自己更好用，所以插件类型的 harness 才会更受欢迎，AI方向一直在变，新的MODEL 最好就不要套用旧版的SKILL跟负担，后面需要在慢慢增加就好，省钱又省时间</p>
]]></description><link>https://lcz.me/post/15512</link><guid isPermaLink="true">https://lcz.me/post/15512</guid><dc:creator><![CDATA[Botio Kuo]]></dc:creator><pubDate>Wed, 02 Sep 2026 12:30:21 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 12:02:12 GMT]]></title><description><![CDATA[<p dir="auto">关于推理，我可以负责任地说，如果你使用DSH，在编程等日常任务中，没啥差距，最起码我什么推理等级都使用过，开了推理除了增加一堆输出，毫无卵用。直接裸模型输出，把harness交给dsh去做。</p>
]]></description><link>https://lcz.me/post/15506</link><guid isPermaLink="true">https://lcz.me/post/15506</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Wed, 02 Sep 2026 12:02:12 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 11:38:48 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/kop-wang" aria-label="Profile: kop-wang">@<bdi>kop-wang</bdi></a></p>
<p dir="auto">沒想到medium and low 只差距一個點</p>
<p dir="auto"><img src="https://upload.lcz.me/uploads/3467bc1d-9eb7-41cb-956f-8eb559ef9d10.jpeg" alt="df7d584e-3f0e-4dc3-a1bb-bbf389259955-image.jpeg" class=" img-fluid img-markdown" /></p>
]]></description><link>https://lcz.me/post/15505</link><guid isPermaLink="true">https://lcz.me/post/15505</guid><dc:creator><![CDATA[kos or]]></dc:creator><pubDate>Wed, 02 Sep 2026 11:38:48 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 11:27:11 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/neo" aria-label="Profile: neo">@<bdi>neo</bdi></a> 有一说一 还是要看你拿来的用途 跟 搭配的 harness , 差异确实不小。。。workflow 跟 一些边界没写好，推理再久还是错</p>
]]></description><link>https://lcz.me/post/15503</link><guid isPermaLink="true">https://lcz.me/post/15503</guid><dc:creator><![CDATA[Botio Kuo]]></dc:creator><pubDate>Wed, 02 Sep 2026 11:27:11 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 11:16:34 GMT]]></title><description><![CDATA[<p dir="auto">关掉推理，连一些基本的测试题都过不了的，稍微复杂点的工具调用也会出错。</p>
]]></description><link>https://lcz.me/post/15498</link><guid isPermaLink="true">https://lcz.me/post/15498</guid><dc:creator><![CDATA[neo]]></dc:creator><pubDate>Wed, 02 Sep 2026 11:16:34 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 09:41:11 GMT]]></title><description><![CDATA[<p dir="auto">所以low是甜蜜點啊....low -&gt; non 可以節省多少時間?</p>
]]></description><link>https://lcz.me/post/15486</link><guid isPermaLink="true">https://lcz.me/post/15486</guid><dc:creator><![CDATA[David Chen]]></dc:creator><pubDate>Wed, 02 Sep 2026 09:41:11 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 08:57:57 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/kop-wang" aria-label="Profile: kop-wang">@<bdi>kop-wang</bdi></a> 是的，但我现在实作的过程中，我不觉得那个分数值得更多的时间。。。 效果真的差不多，这MODEL 思考时间真的太长了。。。</p>
]]></description><link>https://lcz.me/post/15468</link><guid isPermaLink="true">https://lcz.me/post/15468</guid><dc:creator><![CDATA[Botio Kuo]]></dc:creator><pubDate>Wed, 02 Sep 2026 08:57:57 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B 关掉推理～效率快太多 成果差不多 on Wed, 02 Sep 2026 08:38:42 GMT]]></title><description><![CDATA[<p dir="auto"><a href="https://artificialanalysis.ai/models/qwen3-8-27b-non-reasoning" rel="nofollow ugc">https://artificialanalysis.ai/models/qwen3-8-27b-non-reasoning</a></p>
<p dir="auto">这是qwen3.8-27B的关闭思考评测。能力下降的还是比较厉害的。<br />
<img src="https://upload.lcz.me/uploads/2c14ecf8-bdc0-4175-a7ff-1de99e3acd01.jpeg" alt="46c48de9-9f43-4587-b648-fb132bca9017-image.jpeg" class=" img-fluid img-markdown" /></p>
]]></description><link>https://lcz.me/post/15465</link><guid isPermaLink="true">https://lcz.me/post/15465</guid><dc:creator><![CDATA[kop wang]]></dc:creator><pubDate>Wed, 02 Sep 2026 08:38:42 GMT</pubDate></item></channel></rss>