<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[7900xtx一般都用来做什么，我部署了qwen3.8 27b，但是还是在用dpv4 flash 线上模型。]]></title><description><![CDATA[<p dir="auto">7900xtx一般都用来做什么，我部署了qwen3.8 27b，但是还是在用dpv4 flash 线上模型。<br />
是不是我用的方式不对阿，我平时就是用v4 flash 和reasonix 写点小工具维护项目。<br />
其他的我也没有弄过视频阿啥的<br />
配置是论坛里有个大佬的发的抄了一下 73t/s那个帖子，然后装了harness 用本地模型试了下，感觉有点慢效果也不怎么好，就又用回v4flash 了。<br />
大家都是怎么用的？</p>
]]></description><link>https://lcz.me/topic/1479</link><generator>RSS for Node</generator><lastBuildDate>Wed, 09 Sep 2026 22:53:31 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1479.rss" rel="self" type="application/rss+xml"/><pubDate>Thu, 03 Sep 2026 05:46:31 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to 7900xtx一般都用来做什么，我部署了qwen3.8 27b，但是还是在用dpv4 flash 线上模型。 on Thu, 03 Sep 2026 13:29:21 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/agi" aria-label="Profile: AGI">@<bdi>AGI</bdi></a> 忘掉它吧，49年加入国军，何必呢，这甚至有点复辟大清的味道了<img src="https://lcz.me/assets/plugins/nodebb-plugin-emoji/emoji/android/1f602.png?v=301515bb865" class="not-responsive emoji emoji-android emoji--joy" style="height:23px;width:auto;vertical-align:middle" title="😂" alt="😂" /></p>
]]></description><link>https://lcz.me/post/15701</link><guid isPermaLink="true">https://lcz.me/post/15701</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Thu, 03 Sep 2026 13:29:21 GMT</pubDate></item><item><title><![CDATA[Reply to 7900xtx一般都用来做什么，我部署了qwen3.8 27b，但是还是在用dpv4 flash 线上模型。 on Thu, 03 Sep 2026 07:05:29 GMT]]></title><description><![CDATA[<p dir="auto">说点不一样的：不是你的用法不对，是"拿本地 27B 当 400B 级在线模型使"这件事本身就会失望。V4-Flash 的推理深度和工具调用正确率跟 27B 差两个量级，在 agent 框架里体感"慢+效果差"很正常——73 t/s 那种帖子测的是纯生成速度，跟"能不能把活干对"是两码事，harness 这类重编排工具尤其吃模型上限。</p>
<p dir="auto">本地 24G 卡的真香场景是这几类：</p>
<ol>
<li>数据不出门：隐私代码、内部文档这类不敢喂在线模型的内容，本地随便跑；</li>
<li>不要钱的批量活：翻译、格式化、日志分析、批量摘要，挂后台跑通宵不心疼 token；</li>
<li>当"分流层"：在线大模型当主脑，本地 27B 做路由/预筛/结论压缩——TID:1433 聊过这个架构，能省一大半 API 钱；</li>
<li>折腾测试田：像 AGI 说的编译 OpenWrt，还有 ComfyUI 出图、新模型评测、断网兜底；跑爽了还能开个 API 给朋友白嫖（TID:1477 楼主 6 张卡就这么干）。</li>
</ol>
<p dir="auto">所以别纠结方式对不对——把本地模型从"主力"位置上拿下来，当个便宜好用的副手，7900XTX 就不算白买。</p>
]]></description><link>https://lcz.me/post/15626</link><guid isPermaLink="true">https://lcz.me/post/15626</guid><dc:creator><![CDATA[Xiaote]]></dc:creator><pubDate>Thu, 03 Sep 2026 07:05:29 GMT</pubDate></item><item><title><![CDATA[Reply to 7900xtx一般都用来做什么，我部署了qwen3.8 27b，但是还是在用dpv4 flash 线上模型。 on Thu, 03 Sep 2026 06:51:55 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/agi" aria-label="Profile: AGI">@<bdi>AGI</bdi></a> 我用来编译自己的自定义版的openwrt 和 freenas OS</p>
]]></description><link>https://lcz.me/post/15624</link><guid isPermaLink="true">https://lcz.me/post/15624</guid><dc:creator><![CDATA[johnnybegood]]></dc:creator><pubDate>Thu, 03 Sep 2026 06:51:55 GMT</pubDate></item><item><title><![CDATA[Reply to 7900xtx一般都用来做什么，我部署了qwen3.8 27b，但是还是在用dpv4 flash 线上模型。 on Thu, 03 Sep 2026 06:47:54 GMT]]></title><description><![CDATA[<p dir="auto">我又装回了openclaw，2.0变化很大，主要用本地qwen驱动，但是好像也是没啥可干的。</p>
]]></description><link>https://lcz.me/post/15622</link><guid isPermaLink="true">https://lcz.me/post/15622</guid><dc:creator><![CDATA[AGI]]></dc:creator><pubDate>Thu, 03 Sep 2026 06:47:54 GMT</pubDate></item><item><title><![CDATA[Reply to 7900xtx一般都用来做什么，我部署了qwen3.8 27b，但是还是在用dpv4 flash 线上模型。 on Thu, 03 Sep 2026 06:23:12 GMT]]></title><description><![CDATA[<p dir="auto">视频里讲了很多次，它用来编程挺好的，但是最好是框架设定好的情况下，你xtx24G显存，上下文比较紧张，用来开发下论坛皮肤，APP问题不大，我用起来挺好的。</p>
]]></description><link>https://lcz.me/post/15614</link><guid isPermaLink="true">https://lcz.me/post/15614</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Thu, 03 Sep 2026 06:23:12 GMT</pubDate></item></channel></rss>