<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平]]></title><description><![CDATA[<p dir="auto">刚出炉的新鲜评分，Qwen3.8 27B 52 分，与 DSV4 Flash 0731 持平</p>
<p dir="auto">不给上传图片了？大家自己看吧。</p>
<p dir="auto"><a href="https://artificialanalysis.ai/?models=qwen3-6-27b-non-reasoning%2Cqwen3-8-max%2Cdeepseek-v4-pro%2Cqwen3-8-27b%2Cclaude-opus-5%2Cdeepseek-v4-pro-0424%2Cdeepseek-v4-flash-0420%2Cdeepseek-v4-flash%2Cglm-5-2#intelligence-tabs" rel="nofollow ugc">Artificial Analysis Intelligence Index</a></p>
<p dir="auto">当然，评分对应的应该是参数无量化、kv cache 无量化的官方版本。</p>
<p dir="auto">个人体感：3.8 更具主动性，做长任务的能力强了很多。分数或许略有水分，但比 DSV4 Flash Preview 好用是肯定的。和API 相比其实也就是速度慢了点，完全可以替代很多 API 模型了。</p>
<p dir="auto">从 3.6 到 3.8 也才半年，官方已经敢拿它和 Opus 4.6 对比，而不久前 Opus 4.6 还是最好的编程模型。未来应该会超乎很多人想象。</p>
]]></description><link>https://lcz.me/topic/1169</link><generator>RSS for Node</generator><lastBuildDate>Thu, 10 Sep 2026 02:37:41 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1169.rss" rel="self" type="application/rss+xml"/><pubDate>Mon, 17 Aug 2026 18:59:01 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 03:52:11 GMT]]></title><description><![CDATA[<p dir="auto">感谢提醒，大概看了下，还是老问题，小模型的认知水平太差了。<br />
<img src="https://upload.lcz.me/uploads/560bf4f7-9039-4de9-b5d6-ab60b7d67334.jpeg" alt="64ac9893-131f-4412-aa3f-2499f4ec83eb-image.jpeg" class=" img-fluid img-markdown" /></p>
<p dir="auto">这就导致即便他的逻辑能力，长上下文推理能力不弱于flash-preview，甚至能和0731来回。但因为底层认知有巨大鸿沟，所以他更依赖外部数据源的质量。也会比v4-flash-0731有更多的loop和试错。知识量的差距就是这样。</p>
<p dir="auto">如果以后小模型能再进一步，可能小模型+配套行业的数据源的套餐组合是一套新打法。</p>
<p dir="auto">但话又说回来，即便如此，他的实际解决问题的能力依然出众，确实厉害。<br />
<img src="https://upload.lcz.me/uploads/7766fdb4-08eb-4231-bb71-b6a2debee3af.jpeg" alt="43626d02-420b-4f46-a6e7-338298d8f067-image.jpeg" class=" img-fluid img-markdown" /></p>
<p dir="auto">对应的，在如此强的工具、Agent能力之下，256K的上下文其实已经有点瓶颈了。<br />
上重一点的Agent就很有可能多次压缩，导致信息损失。</p>
]]></description><link>https://lcz.me/post/12653</link><guid isPermaLink="true">https://lcz.me/post/12653</guid><dc:creator><![CDATA[kop wang]]></dc:creator><pubDate>Tue, 18 Aug 2026 03:52:11 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 03:36:21 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E4%B9%9D%E9%BE%99%E6%9D%A8%E7%94%9F" aria-label="Profile: 九龙杨生">@<bdi>九龙杨生</bdi></a> opencode Go 的负责人多次表示，目前还没人能在和原厂能力持平的前提下，总成本低于原厂。</p>
<p dir="auto">供参考</p>
]]></description><link>https://lcz.me/post/12650</link><guid isPermaLink="true">https://lcz.me/post/12650</guid><dc:creator><![CDATA[kop wang]]></dc:creator><pubDate>Tue, 18 Aug 2026 03:36:21 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 02:40:40 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E4%B9%9D%E9%BE%99%E6%9D%A8%E7%94%9F" aria-label="Profile: 九龙杨生">@<bdi>九龙杨生</bdi></a></p>
<p dir="auto">Deepseek-V4-Flash 我沒開過Max, 都用一般的<br />
但根據“某次測試”比較 GPT5.6-Luna Max 是比Terra Medium強</p>
<p dir="auto">长链任务可能也要看 Agent 框架的Harness 能力 兩者配合就適合跑长链任务, Codex 接上Qwen3.8 27B 或許可以測試一番<br />
Codex vs. Deepseek harness 框架</p>
]]></description><link>https://lcz.me/post/12639</link><guid isPermaLink="true">https://lcz.me/post/12639</guid><dc:creator><![CDATA[kos or]]></dc:creator><pubDate>Tue, 18 Aug 2026 02:40:40 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 02:15:22 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/kos-or" aria-label="Profile: kos-or">@<bdi>kos-or</bdi></a> V4 FLASH MAX这么强的吗？和PRO基本上就差不多了；<br />
Qwen3.8 27B估计就是输在长链任务上</p>
]]></description><link>https://lcz.me/post/12638</link><guid isPermaLink="true">https://lcz.me/post/12638</guid><dc:creator><![CDATA[九龙杨生]]></dc:creator><pubDate>Tue, 18 Aug 2026 02:15:22 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 02:34:33 GMT]]></title><description><![CDATA[<p dir="auto">加上了GPT5.6-SOL and Luna Max 比較<br />
我再找機會讓Luna Max and Qwen3.8-27 實際對壘</p>
<p dir="auto"><img src="https://upload.lcz.me/uploads/b817b53a-ee7a-4514-a240-3974fd42913f.jpeg" alt="5d9bf3cd-dc10-402b-9668-ebb13a2e3dc1-image.jpeg" class=" img-fluid img-markdown" /></p>
]]></description><link>https://lcz.me/post/12636</link><guid isPermaLink="true">https://lcz.me/post/12636</guid><dc:creator><![CDATA[kos or]]></dc:creator><pubDate>Tue, 18 Aug 2026 02:34:33 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 01:19:22 GMT]]></title><description><![CDATA[<p dir="auto">感觉不在一个赛道，27B的8位权重 ，8位KV CACHE，48 64G 72G显存应该都可以流畅的跑256K上下文了。 v4 flash要流畅本地部署，至少还要在单CPU上挂 128G 或256G DDR5内存吧？</p>
]]></description><link>https://lcz.me/post/12632</link><guid isPermaLink="true">https://lcz.me/post/12632</guid><dc:creator><![CDATA[stxpnet]]></dc:creator><pubDate>Tue, 18 Aug 2026 01:19:22 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 00:35:35 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E4%B9%9D%E9%BE%99%E6%9D%A8%E7%94%9F" aria-label="Profile: 九龙杨生">@<bdi>九龙杨生</bdi></a> 缓存和成本很难做到官方那样，其它的没啥区别。只要你买到的便宜，你何必在乎呢？只不过第三方的V4 flash很难和官方的比速度，因为Engram架构，HCA/CSA这些注意力机制是DeepSeek的独门技术，没有开源。你自己用下要是没问题就搞就是了。自己用没必要啊，一共屁大需求，还不如把qwen3.8 27b玩明白。它能干脏活，剩下的给在线的。</p>
]]></description><link>https://lcz.me/post/12628</link><guid isPermaLink="true">https://lcz.me/post/12628</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Tue, 18 Aug 2026 00:35:35 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 00:13:54 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/terry" aria-label="Profile: terry">@<bdi>terry</bdi></a> 有两个渠道，一个是移动自建的，一个是火山的，我主要就是担心模型能力不如官方；渠道上游提供商没问题，只是可能有中间商赚差价，不过最终只要能够给出有竞争力的价格就行；</p>
]]></description><link>https://lcz.me/post/12627</link><guid isPermaLink="true">https://lcz.me/post/12627</guid><dc:creator><![CDATA[九龙杨生]]></dc:creator><pubDate>Tue, 18 Aug 2026 00:13:54 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 00:05:40 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E4%B9%9D%E9%BE%99%E6%9D%A8%E7%94%9F" aria-label="Profile: 九龙杨生">@<bdi>九龙杨生</bdi></a> 你要是做生意没啥，有机会就赚，火山引擎是字节跳动的，大厂实力不用担心。问题是你接触的确实是火山吗？你要是自己用，我擦，这三瓜俩枣的，何必呢。</p>
]]></description><link>https://lcz.me/post/12626</link><guid isPermaLink="true">https://lcz.me/post/12626</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Tue, 18 Aug 2026 00:05:40 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Tue, 18 Aug 2026 00:01:18 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/terry" aria-label="Profile: terry">@<bdi>terry</bdi></a> 这方面我就不专业了，最近有人跟我对接火山方舟的DeepSeek API资源，能够拿到折扣价；他这种抖音下面自建的和官方的模型能力是不是一样的呢？还是说在什么地方存在性能差距？</p>
]]></description><link>https://lcz.me/post/12625</link><guid isPermaLink="true">https://lcz.me/post/12625</guid><dc:creator><![CDATA[九龙杨生]]></dc:creator><pubDate>Tue, 18 Aug 2026 00:01:18 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Mon, 17 Aug 2026 23:51:47 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E4%B9%9D%E9%BE%99%E6%9D%A8%E7%94%9F" aria-label="Profile: 九龙杨生">@<bdi>九龙杨生</bdi></a> 最主要是注意力机制，上下文可以压缩，但是召回准确率差别是致命的。</p>
]]></description><link>https://lcz.me/post/12624</link><guid isPermaLink="true">https://lcz.me/post/12624</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Mon, 17 Aug 2026 23:51:47 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Mon, 17 Aug 2026 23:48:21 GMT]]></title><description><![CDATA[<p dir="auto">感觉还是上下文长度限制了本地模型的能力；压缩几次就记不住前面的了。</p>
]]></description><link>https://lcz.me/post/12623</link><guid isPermaLink="true">https://lcz.me/post/12623</guid><dc:creator><![CDATA[九龙杨生]]></dc:creator><pubDate>Mon, 17 Aug 2026 23:48:21 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Mon, 17 Aug 2026 22:07:12 GMT]]></title><description><![CDATA[<p dir="auto">1，并没有超越DeepSeek，我也非常看好Qwen3.8 27b，而且我A卡 N卡都部署了，今天的视频大家可以来看。单轮编码能力很好，但是它长线还是完全不如V4 Flash，基础能力相差很大，迭代之后差距更大。只能说小玩玩，后续想办法设计下特殊的工作流可以。超越V4 Flash的说法是碰瓷了。跑分都是轻任务，甚至单论，意义不大。而且Qwen开了思考之后，速度慢成狗，不开思考会降智力，那才是Qwen的真实水平。</p>
<p dir="auto">2，应该是你的问题，没人限制你，大家都能上传，你换个浏览器看看。</p>
]]></description><link>https://lcz.me/post/12614</link><guid isPermaLink="true">https://lcz.me/post/12614</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Mon, 17 Aug 2026 22:07:12 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Mon, 17 Aug 2026 20:44:36 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E5%83%8F%E7%B4%A0%E7%9B%92%E5%AD%90" aria-label="Profile: 像素盒子">@<bdi>像素盒子</bdi></a>  "错误 您没有权限执行此操作。" 一张 23KB 的 png，转成更小的 jpg 也不行。</p>
]]></description><link>https://lcz.me/post/12609</link><guid isPermaLink="true">https://lcz.me/post/12609</guid><dc:creator><![CDATA[stakira]]></dc:creator><pubDate>Mon, 17 Aug 2026 20:44:36 GMT</pubDate></item><item><title><![CDATA[Reply to Qwen3.8 27B AA综合评分与 DeepSeek V4 Flash 持平 on Mon, 17 Aug 2026 20:21:48 GMT]]></title><description><![CDATA[<p dir="auto">能上传，图片不要太大！</p>
]]></description><link>https://lcz.me/post/12608</link><guid isPermaLink="true">https://lcz.me/post/12608</guid><dc:creator><![CDATA[像素盒子]]></dc:creator><pubDate>Mon, 17 Aug 2026 20:21:48 GMT</pubDate></item></channel></rss>