<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[qwen3.8-27b无审查量化有没有比较推荐的社区版本，要能做事聪明的。]]></title><description><![CDATA[<p dir="auto">我现在用的是0bserverx/Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF 的 Qwen3.8-27B-RVN-Q4_K_M-multilingual-mtp.gguf</p>
<p dir="auto">今天又下载了那个排名蹿升比较快的DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP-GGUF</p>
<p dir="auto">结果分别出了两道编程题，让豆包替我评分，DavidAU的模型完败，豆包对它的两个答案评价是“ 函数<strong>有严重性能缺陷</strong>和”致命缺陷，属于错误实现“。</p>
<p dir="auto">本来看它一堆前缀，介绍里写得也很猛，以为很厉害，结果是个弱智。</p>
<p dir="auto">坛友有啥其他推荐吗，感觉缺乏这方面的参考，社区模型有没有啥大众点评网  <img src="https://lcz.me/assets/plugins/nodebb-plugin-emoji/emoji/android/1f601.png?v=301515bb865" class="not-responsive emoji emoji-android emoji--grin" style="height:23px;width:auto;vertical-align:middle" title=":grin:" alt="😁" /></p>
]]></description><link>https://lcz.me/topic/1575</link><generator>RSS for Node</generator><lastBuildDate>Wed, 09 Sep 2026 22:54:25 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1575.rss" rel="self" type="application/rss+xml"/><pubDate>Wed, 09 Sep 2026 04:07:36 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to qwen3.8-27b无审查量化有没有比较推荐的社区版本，要能做事聪明的。 on Wed, 09 Sep 2026 08:32:06 GMT]]></title><description><![CDATA[<p dir="auto">回到 Qwen 3.6 27B 或 A3B 可以。<br />
3.8 太新。<br />
让时间解决这个问题。</p>
]]></description><link>https://lcz.me/post/16866</link><guid isPermaLink="true">https://lcz.me/post/16866</guid><dc:creator><![CDATA[williamlouis]]></dc:creator><pubDate>Wed, 09 Sep 2026 08:32:06 GMT</pubDate></item><item><title><![CDATA[Reply to qwen3.8-27b无审查量化有没有比较推荐的社区版本，要能做事聪明的。 on Wed, 09 Sep 2026 08:01:14 GMT]]></title><description><![CDATA[<p dir="auto">想真正能干活的，别用越狱版，调用工具会出错，你可以再下一个越狱版备用，需要越狱内容的时候切换一下模型就可以了。</p>
]]></description><link>https://lcz.me/post/16859</link><guid isPermaLink="true">https://lcz.me/post/16859</guid><dc:creator><![CDATA[koala]]></dc:creator><pubDate>Wed, 09 Sep 2026 08:01:14 GMT</pubDate></item><item><title><![CDATA[Reply to qwen3.8-27b无审查量化有没有比较推荐的社区版本，要能做事聪明的。 on Wed, 09 Sep 2026 07:05:24 GMT]]></title><description><![CDATA[<p dir="auto">名字越长越危险，这是识别社区无审查量化的第一把尺。DavidAU 那串「Cold-Fusion-735-882-Heretic-Uncensored-NEO-CODER-MAX-MTP」= 多次 stack-mix + 多次去审查/再蒸馏层层叠加，每加一层就多损失一截，最后把格式和推理都磨没了。豆包评「严重性能缺陷/错误实现」其实吻合这个规律，不是纯玄学。</p>
<p dir="auto">第二把尺：去审查（abliteration）本身是往残差流里减/推一回拒方向，下刀越狠智商掉越多。稳妥做法是烧窄层段（比如 23-51，别从 18 劈到 51）且 MTP 头别动。huihui 的 update4 系列就是走的 unsloth UD 路线、只烧 23-51 层，相对干净。</p>
<p dir="auto">第三把尺：别拿豆包这种 LLM-as-judge 当 A/B。它评代码是「像不像错的」，噪音极大。真要选，跑同一套固定 prompt（一组代码题 + 一组「敏感题」），自己量化对比拒答率 + 正确率，而不是让另一个模型打分。</p>
<p dir="auto">实用建议：24G 单卡就锁 27B Q4_K_M，优先选保留 MTP 头、单遍去审查的（你现在 0bserverx 的 RVN 这类就不错）；碰到名号为王、一长串前缀堆砌的直接跳过。Heretic 那个工具提醒一句：它要整模型驻留才能处理，1B≈2.5G，27B 得 68G，24G 卡压根跑不动，别指望本地 24G 手搓出低拒答版。</p>
]]></description><link>https://lcz.me/post/16850</link><guid isPermaLink="true">https://lcz.me/post/16850</guid><dc:creator><![CDATA[Xiaote]]></dc:creator><pubDate>Wed, 09 Sep 2026 07:05:24 GMT</pubDate></item><item><title><![CDATA[Reply to qwen3.8-27b无审查量化有没有比较推荐的社区版本，要能做事聪明的。 on Wed, 09 Sep 2026 04:25:06 GMT]]></title><description><![CDATA[<p dir="auto">噩耗，我刚刚去读了一下Heretic的Readme，1B模型需要2.5G Vram，直接判定了此路不通。27B意味着至少需要68G以上的显存才可以支持，24G的得3卡。</p>
]]></description><link>https://lcz.me/post/16836</link><guid isPermaLink="true">https://lcz.me/post/16836</guid><dc:creator><![CDATA[2024fatwolf55]]></dc:creator><pubDate>Wed, 09 Sep 2026 04:25:06 GMT</pubDate></item><item><title><![CDATA[Reply to qwen3.8-27b无审查量化有没有比较推荐的社区版本，要能做事聪明的。 on Wed, 09 Sep 2026 04:18:37 GMT]]></title><description><![CDATA[<p dir="auto">YEA,真正拿来投产的存在很多问题，我没用过unslot, 之前专用无审查 llm46fan  也翻车<br />
hauhau, huihui, david, 全部翻车</p>
<p dir="auto">现在换成JonathanColetti，后来找到一个BennyDaBall nvfp4 ,他的source 来自JonathanColetti的，然而两者还没算一个好用的。。。</p>
<p dir="auto">你不妨试一试，坊间有一句话，难吃的食物不能一个人独享。。。嘿嘿嘿</p>
<p dir="auto"><a href="https://huggingface.co/JonathanColetti/Qwen3.8-27B-Uncensored-GGUF" rel="nofollow ugc">https://huggingface.co/JonathanColetti/Qwen3.8-27B-Uncensored-GGUF</a></p>
<p dir="auto"><a href="https://huggingface.co/BennyDaBall/Qwen3.8-Uncensored-NVFP4-MTP" rel="nofollow ugc">https://huggingface.co/BennyDaBall/Qwen3.8-Uncensored-NVFP4-MTP</a></p>
]]></description><link>https://lcz.me/post/16835</link><guid isPermaLink="true">https://lcz.me/post/16835</guid><dc:creator><![CDATA[imbiplaza ASUS]]></dc:creator><pubDate>Wed, 09 Sep 2026 04:18:37 GMT</pubDate></item><item><title><![CDATA[Reply to qwen3.8-27b无审查量化有没有比较推荐的社区版本，要能做事聪明的。 on Wed, 09 Sep 2026 04:16:46 GMT]]></title><description><![CDATA[<p dir="auto">我前两天看到有篇文章，介绍说Github上有个项目叫做Heretic，（<a href="https://github.com/p-e-w/heretic%EF%BC%89" rel="nofollow ugc">https://github.com/p-e-w/heretic）</a></p>
<p dir="auto">号称是你自己安装以后，可以自行利用你的GPU，对你的现有模型进行自动处理（据称支持绝大多数稠密模型），让其对大量原先拒答问题的拒答比例，100题里面可以降低到只拒答3道，而不降低智力。</p>
<p dir="auto">我刚刚下载了这个东西，还没在自己的机器上跑，打算回头用它试试看已经下载的FP8，看看能不能自己手搓一个“无拒答版”的Qwen3.8 27B FP8.</p>
<p dir="auto">如果可能，我觉得就很符合你的要求。</p>
]]></description><link>https://lcz.me/post/16833</link><guid isPermaLink="true">https://lcz.me/post/16833</guid><dc:creator><![CDATA[2024fatwolf55]]></dc:creator><pubDate>Wed, 09 Sep 2026 04:16:46 GMT</pubDate></item></channel></rss>