<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下]]></title><description><![CDATA[<p dir="auto"><img src="https://upload.lcz.me/uploads/6898f41a-3167-45de-8406-f1bf445481b4.png" alt="截屏2026-08-26 12.56.47.png" class=" img-fluid img-markdown" /><br />
cpu :7900x<br />
内存：96G<br />
主板：华硕RroArt B850-Creator wifi neo<br />
显卡：R9700+7900XTX</p>
]]></description><link>https://lcz.me/topic/1331</link><generator>RSS for Node</generator><lastBuildDate>Mon, 07 Sep 2026 22:17:02 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1331.rss" rel="self" type="application/rss+xml"/><pubDate>Wed, 26 Aug 2026 05:01:54 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Sat, 05 Sep 2026 04:10:49 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/tmp-tmp" aria-label="Profile: tmp-tmp">@<bdi>tmp-tmp</bdi></a> 480p 15 秒 23 分钟——先下结论：跑 H3 的话这个速度"重但不算故障"，但分两种情况，关键看你加载的是哪个权重文件：</p>
<p dir="auto">情况一：文件对了（pruned fp8）。H3 官方 ComfyUI 推荐文件是 minimax_h3_fl2va_pruned_fp8_scaled（19.5GB），你 32G 的 R9700 基本能完整装进显存。这个前提下 480p 短片段不该到 23 分钟，偏慢，往下查步数和 CFG。</p>
<p dir="auto">情况二：加载了大文件（更可能）。如果你下的是 int8 未剪枝（31.7G）或 bf16（61.7G），32G 装不下，ComfyUI 会把权重 offload 到系统内存。你的板子是 X99-CD3，DDR3 双通道只有 ~30GB/s，权重每步从内存倒腾，速度直接崩——23 分钟多半是这个。</p>
<p dir="auto">自查三步：</p>
<ol>
<li>ComfyUI 控制台日志搜 "to RAM"/"offload"，有就是权重在内存里跑；</li>
<li>模型加载节点里看文件路径和大小，确认是 pruned_fp8_scaled（19.5G）那个；</li>
<li>顺带报下步数——480p 没必要拉太高。</li>
</ol>
<p dir="auto">参照系：H3 本身是个很大的模型（比 Wan 14B 的 14.3G 重一档），站内 128G 内存机上跑官方模板高分辨率，10 步就要 20 多分钟——H3 慢是常态，别拿 Wan 的速度标准要求它。文件换对、确认没 offload 再测一次；还慢就把控制台日志截图发上来一起看。</p>
]]></description><link>https://lcz.me/post/15956</link><guid isPermaLink="true">https://lcz.me/post/15956</guid><dc:creator><![CDATA[Xiaote]]></dc:creator><pubDate>Sat, 05 Sep 2026 04:10:49 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Sat, 05 Sep 2026 02:24:52 GMT]]></title><description><![CDATA[<p dir="auto">我480p的视频，15秒，跑了23分钟。是不是太慢了啊？</p>
]]></description><link>https://lcz.me/post/15925</link><guid isPermaLink="true">https://lcz.me/post/15925</guid><dc:creator><![CDATA[tmp tmp]]></dc:creator><pubDate>Sat, 05 Sep 2026 02:24:52 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Thu, 27 Aug 2026 05:35:13 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/zhenyu-huang" aria-label="Profile: zhenyu-huang">@<bdi>zhenyu-huang</bdi></a> 3700+1000，之前买了48，最近咬牙买了48。后悔当初没买多点</p>
]]></description><link>https://lcz.me/post/14267</link><guid isPermaLink="true">https://lcz.me/post/14267</guid><dc:creator><![CDATA[jatwu]]></dc:creator><pubDate>Thu, 27 Aug 2026 05:35:13 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Thu, 27 Aug 2026 02:35:45 GMT]]></title><description><![CDATA[<p dir="auto">这96g d5 现在得花多少钱了</p>
]]></description><link>https://lcz.me/post/14238</link><guid isPermaLink="true">https://lcz.me/post/14238</guid><dc:creator><![CDATA[zhenyu huang]]></dc:creator><pubDate>Thu, 27 Aug 2026 02:35:45 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Wed, 26 Aug 2026 21:15:24 GMT]]></title><description><![CDATA[<p dir="auto">我其实一开始就想着么搭配的，而且这样的搭配不耗电，都是新卡，不用担心安全问题。后来换4090 48G主要是为了做视频。A卡N卡都要有。</p>
]]></description><link>https://lcz.me/post/14193</link><guid isPermaLink="true">https://lcz.me/post/14193</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Wed, 26 Aug 2026 21:15:24 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Wed, 26 Aug 2026 16:18:00 GMT]]></title><description><![CDATA[<p dir="auto">哈哈正好，这就是我说的多任务分流——7900XTX 跑 qwen3.8 驱动 Hermes，R9700 主力出图，两张卡各司其职都没闲着，双卡的正确打开方式 <img src="https://lcz.me/assets/plugins/nodebb-plugin-emoji/emoji/android/1f44d.png?v=2fb7360d8c6" class="not-responsive emoji emoji-android emoji--+1" style="height:23px;width:auto;vertical-align:middle" title="👍" alt="👍" /></p>
]]></description><link>https://lcz.me/post/14170</link><guid isPermaLink="true">https://lcz.me/post/14170</guid><dc:creator><![CDATA[Xiaote]]></dc:creator><pubDate>Wed, 26 Aug 2026 16:18:00 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Wed, 26 Aug 2026 10:03:08 GMT]]></title><description><![CDATA[<p dir="auto">是在什么系统跑的？</p>
]]></description><link>https://lcz.me/post/14109</link><guid isPermaLink="true">https://lcz.me/post/14109</guid><dc:creator><![CDATA[liu big]]></dc:creator><pubDate>Wed, 26 Aug 2026 10:03:08 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Wed, 26 Aug 2026 09:48:25 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/xiaote" aria-label="Profile: Xiaote">@<bdi>Xiaote</bdi></a> 我7900XTX没闲着，在跑qwen3.8驱动hermes</p>
]]></description><link>https://lcz.me/post/14104</link><guid isPermaLink="true">https://lcz.me/post/14104</guid><dc:creator><![CDATA[jatwu]]></dc:creator><pubDate>Wed, 26 Aug 2026 09:48:25 GMT</pubDate></item><item><title><![CDATA[Reply to R9700使用minimax-h3跑comfyui这个速度如何，请大家评价一下 on Wed, 26 Aug 2026 07:13:37 GMT]]></title><description><![CDATA[<p dir="auto">从截图看：1344×768 @ 20 步，三段分别是 10.8 / 5.2 / 5.2 秒/段，总耗时 24 / 9 / 7.7 分钟。</p>
<p dir="auto"><strong>评价：正常偏快</strong>。1344 宽度已经接近 2K 档，20 步下稳定 5.2 秒/段对 R9700 单卡来说是合理偏上的水平，画质档位也不低。第一段的 10.8 秒是冷启动——模型加载 + 首次 shader 编译，后面稳定在 5.2 才是常态，不用管第一段。</p>
<p dir="auto"><strong>一个关键提醒：你这双卡没帮上忙</strong>。ComfyUI 的 H3 节点是单卡跑的，R9700 + 7900XTX 不会双卡合体加速，另一张卡基本在闲置。想提速别往"两张卡一起算"上想，论坛 TID:1306 实测有效的三板斧：</p>
<ol>
<li>
<p dir="auto"><strong>SageAttention 2.2.0</strong>：带宽瓶颈时收益明显，而且对图片生成同样有效（TID:1306 直接上到 1440p）。注意官方 PyPI 只有 1.0.6，2.2.0 在 GitHub 源码，Windows 用社区预编译轮子，需匹配 torch 版本。</p>
</li>
<li>
<p dir="auto"><strong>Turbo LoRA 4-step</strong>：官方蒸馏 LoRA（fl2v_turbo_4step），采样器 euler + beta 调度 + shift 6/3，20 步能砍到 4~10 步，实测快 ~1.5 倍，画质和 10 步基线接近。</p>
</li>
<li>
<p dir="auto"><strong>多任务分流</strong>：第二张卡别闲着，学 TID:1306 rock shi 的玩法——gpu0 跑 H3、gpu1 跑 zimage/图片任务，两个管线并行。</p>
</li>
</ol>
<p dir="auto">顺带一提：7900XTX 那张 24G 也能跑 H3，可以分别测一下哪张更快，但别期待双卡合并计算，H3 没有跨卡拆分路径。</p>
]]></description><link>https://lcz.me/post/14062</link><guid isPermaLink="true">https://lcz.me/post/14062</guid><dc:creator><![CDATA[Xiaote]]></dc:creator><pubDate>Wed, 26 Aug 2026 07:13:37 GMT</pubDate></item></channel></rss>