<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[Qwen3.8-27B-Q4_K_M 極端測驗]]></title><description><![CDATA[<p dir="auto">我使用了兩個SOTA AI出了兩個高難度的考題 來測試<br />
Unsloth Qwen3.8-27-Q4_K_M (KV Cache : Q8_0)<br />
答題後再貼給出題者AI評分, 考題難度太高了 所以我沒研究推理過程, 但我對照答案 兩題都答對</p>
<pre><code>以下是GPT5.6-SOL-High 的評價 : 
這份 Agent 回答非常強。依照我前面訂的 100 分評分規則，
我會給它 98/100，S+ 級。
</code></pre>
<p dir="auto"><img src="https://upload.lcz.me/uploads/ba271210-a6e7-4995-a8cf-f5af67dad705.jpeg" alt="381525a7-c19d-49d8-a4f7-0e4d429b5987-image.jpeg" class=" img-fluid img-markdown" /></p>
<pre><code>以下是Sonnet5-Extra的評價 : 
這份提交完全正確，而且推理路徑在幾個地方比我的 model answer 更精簡優雅。
總分：99 / 100 為什麼扣那 1 分（其實不算真正的瑕疵）
我原本設計的路徑，會在加密收尾階段面對...，需要逐一測試淘汰兩組。
</code></pre>
<p dir="auto"><img src="https://upload.lcz.me/uploads/c3db3965-d388-46f5-8af6-1e2da7793e7c.jpeg" alt="cf113f5f-7dac-4b17-98bd-d590d7f4dac7-image.jpeg" class=" img-fluid img-markdown" /></p>
]]></description><link>https://lcz.me/topic/1133</link><generator>RSS for Node</generator><lastBuildDate>Wed, 09 Sep 2026 22:53:33 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1133.rss" rel="self" type="application/rss+xml"/><pubDate>Sat, 15 Aug 2026 09:14:21 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sun, 16 Aug 2026 06:44:42 GMT]]></title><description><![CDATA[<blockquote>
<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/stxpnet" aria-label="Profile: stxpnet">@<bdi>stxpnet</bdi></a> <a href="/post/12357">said</a>:</p>
<p dir="auto">一是说内部思维链比3.6长</p>
</blockquote>
<p dir="auto">我昨天第一次測試的極限考題, 是前一段時間Gemini Pro 出的<br />
沒想到Gemini Pro自己也出亂子, 本來應該是唯一解, 變成居然有11種解法,<br />
但題目又告訴其他受測AI是唯一解,  Qwen3.8 就一直算這唯一解output 2萬多個tokens , 我就把它關了;</p>
<p dir="auto">最後是GPT5.6-SOL 解出來說不可能有唯一解, 而是有11解; 極限考題最後由SOL幫忙改設計成唯一解</p>
]]></description><link>https://lcz.me/post/12361</link><guid isPermaLink="true">https://lcz.me/post/12361</guid><dc:creator><![CDATA[kos or]]></dc:creator><pubDate>Sun, 16 Aug 2026 06:44:42 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sun, 16 Aug 2026 06:35:46 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/whynot" aria-label="Profile: whynot">@<bdi>whynot</bdi></a></p>
<p dir="auto">這是Reddit上網友的分享 貪食蛇遊戲 很好玩</p>
<p dir="auto"><a href="https://s3.fursov.family/shares/snake3d.html" rel="nofollow ugc">https://s3.fursov.family/shares/snake3d.html</a></p>
<p dir="auto">" slavik-dev•1天前 我有一個問題，沒有一個模型能夠給出令人滿意的結果，只有 Opus 能夠做到。And this model Qwen3.8 (using UD-Q5_K_XL) did it! "</p>
<p dir="auto">prompt :</p>
<pre><code>write snake game on the sphere. The head of the snake it fixed in the center and the sphere is rotating. use HTML, CSS and JavaScript.
The visible part of sphere shall be fully visible in the webView, not partially.
The starting length of the snake shall be 3 and increasing every time the snake hit the food.
Use keyboard control: LEFT and RIGHT arrows.
</code></pre>
<p dir="auto">翻譯後的中文Prompt:</p>
<pre><code>在球體上寫一個貪吃蛇遊戲。蛇頭固定在球體中心，球體旋轉。使用 HTML、CSS 和 JavaScript。
球體的可見部分在 webView 中應完全可見，而不是部分可見。
蛇的初始長度為 3，每次蛇碰到食物後長度都會增加。
使用鍵盤控制：左箭頭和右箭頭。
</code></pre>
]]></description><link>https://lcz.me/post/12360</link><guid isPermaLink="true">https://lcz.me/post/12360</guid><dc:creator><![CDATA[kos or]]></dc:creator><pubDate>Sun, 16 Aug 2026 06:35:46 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sun, 16 Aug 2026 05:51:09 GMT]]></title><description><![CDATA[<p dir="auto">有两个坑，一是说内部思维链比3.6长，另一个是说 要用新的v22的chat template才能适配 3.8</p>
]]></description><link>https://lcz.me/post/12357</link><guid isPermaLink="true">https://lcz.me/post/12357</guid><dc:creator><![CDATA[stxpnet]]></dc:creator><pubDate>Sun, 16 Aug 2026 05:51:09 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sun, 16 Aug 2026 04:47:31 GMT]]></title><description><![CDATA[<p dir="auto">昨天试了3.8 27b,  3.6 27b完美完成的贪吃蛇和坦克大战，居然q4kl,Q5都没有完成，反复修改两次，也没有成功。不过画面的确比3.6好不少，只是无法运行。</p>
]]></description><link>https://lcz.me/post/12354</link><guid isPermaLink="true">https://lcz.me/post/12354</guid><dc:creator><![CDATA[whynot]]></dc:creator><pubDate>Sun, 16 Aug 2026 04:47:31 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 15:54:48 GMT]]></title><description><![CDATA[<blockquote>
<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/vosrock" aria-label="Profile: vosrock">@<bdi>vosrock</bdi></a> <a href="/post/12323">said</a>:</p>
<p dir="auto">我现在用的是THINKCAP的模型，相当于是社区对3。6长期优化版本，体验比3.8要好的多</p>
</blockquote>
<p dir="auto">我還沒測試過agentic use, 等之後看看社區能不能優化,<br />
我對模型的解題能力有興趣, 能應付得了前沿模型的測試 對於27B模型來說 算是蠻厲害的了 不過我還沒用相同的問題測試過其他本地模型</p>
<p dir="auto">測試過程 其中有一題 GPT5.6-Terra-medium 答錯, Qwen3.8答對, GPT-Luna-max 答對</p>
]]></description><link>https://lcz.me/post/12328</link><guid isPermaLink="true">https://lcz.me/post/12328</guid><dc:creator><![CDATA[kos or]]></dc:creator><pubDate>Sat, 15 Aug 2026 15:54:48 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 15:20:58 GMT]]></title><description><![CDATA[<p dir="auto">要想3.8的27B好用，还需要时间</p>
]]></description><link>https://lcz.me/post/12324</link><guid isPermaLink="true">https://lcz.me/post/12324</guid><dc:creator><![CDATA[vosrock]]></dc:creator><pubDate>Sat, 15 Aug 2026 15:20:58 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 15:20:20 GMT]]></title><description><![CDATA[<p dir="auto">讲真，基于Q4 KM量化的模型，3.8没有给到我惊艳的感觉，我现在用的是THINKCAP的模型，相当于是社区对3。6长期优化版本，体验比3.8要好的多</p>
]]></description><link>https://lcz.me/post/12323</link><guid isPermaLink="true">https://lcz.me/post/12323</guid><dc:creator><![CDATA[vosrock]]></dc:creator><pubDate>Sat, 15 Aug 2026 15:20:20 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 13:47:14 GMT]]></title><description><![CDATA[<p dir="auto">看起来不错 赶紧操练起来</p>
]]></description><link>https://lcz.me/post/12316</link><guid isPermaLink="true">https://lcz.me/post/12316</guid><dc:creator><![CDATA[Grayson Ren]]></dc:creator><pubDate>Sat, 15 Aug 2026 13:47:14 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 10:01:00 GMT]]></title><description><![CDATA[<p dir="auto">7900xtx，去升级vulkan驱动。昨晚升级了。Qwen3.8-27B跑最高tg70+ t/s。pp和原来差不多。</p>
]]></description><link>https://lcz.me/post/12289</link><guid isPermaLink="true">https://lcz.me/post/12289</guid><dc:creator><![CDATA[kenshin]]></dc:creator><pubDate>Sat, 15 Aug 2026 10:01:00 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 09:49:27 GMT]]></title><description><![CDATA[<p dir="auto">我更换了之后测试一下，发现它的速度比3.6版的还慢，之前3.6版的token有44 t/s，3.8版的剩下34 t/s</p>
]]></description><link>https://lcz.me/post/12288</link><guid isPermaLink="true">https://lcz.me/post/12288</guid><dc:creator><![CDATA[艷陽天]]></dc:creator><pubDate>Sat, 15 Aug 2026 09:49:27 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 09:36:18 GMT]]></title><description><![CDATA[<p dir="auto">参照我的模型联网帖。联网后就解决了。</p>
]]></description><link>https://lcz.me/post/12284</link><guid isPermaLink="true">https://lcz.me/post/12284</guid><dc:creator><![CDATA[williamlouis]]></dc:creator><pubDate>Sat, 15 Aug 2026 09:36:18 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 09:28:14 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/stxpnet" aria-label="Profile: stxpnet">@<bdi>stxpnet</bdi></a></p>
<p dir="auto">沒呀 我沒有完整的測試框架, 我只能用之前Qwen3.6-27B 挑戰失敗的題目重新挑戰Qwen3.8-27B, 3.8 通過了 7/10難度的考題</p>
<p dir="auto">然後我用GPT5.6-SOL-High 繼續出題加碼到9/10難度的 居然還是被Q4_K_M 精度的模型答對了</p>
]]></description><link>https://lcz.me/post/12283</link><guid isPermaLink="true">https://lcz.me/post/12283</guid><dc:creator><![CDATA[kos or]]></dc:creator><pubDate>Sat, 15 Aug 2026 09:28:14 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8-27B-Q4_K_M 極端測驗 on Sat, 15 Aug 2026 09:21:08 GMT]]></title><description><![CDATA[<p dir="auto">有没有全面一点的a b 测试呢？我昨晚试了感觉千问团队在耍猴，视觉塔的大小一模一样。问它2026年的事情它啥都不知道，说2026年还没来</p>
]]></description><link>https://lcz.me/post/12278</link><guid isPermaLink="true">https://lcz.me/post/12278</guid><dc:creator><![CDATA[stxpnet]]></dc:creator><pubDate>Sat, 15 Aug 2026 09:21:08 GMT</pubDate></item></channel></rss>