<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0"><channel><title><![CDATA[M5 Ultra的1.2T/s高带宽有多大用处？]]></title><description><![CDATA[<p dir="auto">M芯片prefill速度慢如蜗牛 估计就N卡的1/10速度；我在想选配96GB到底能有多大生产力</p>
]]></description><link>https://lcz.me/topic/1489</link><generator>RSS for Node</generator><lastBuildDate>Wed, 09 Sep 2026 21:48:14 GMT</lastBuildDate><atom:link href="https://lcz.me/topic/1489.rss" rel="self" type="application/rss+xml"/><pubDate>Thu, 03 Sep 2026 16:48:30 GMT</pubDate><ttl>60</ttl><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Sun, 06 Sep 2026 03:22:35 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E5%BC%A0%E5%85%89%E7%92%9E" aria-label="Profile: 张光璞">@<bdi>张光璞</bdi></a> 我2080峰值到过7000哈哈哈</p>
]]></description><link>https://lcz.me/post/16111</link><guid isPermaLink="true">https://lcz.me/post/16111</guid><dc:creator><![CDATA[rock shi]]></dc:creator><pubDate>Sun, 06 Sep 2026 03:22:35 GMT</pubDate></item><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Sat, 05 Sep 2026 15:36:20 GMT]]></title><description><![CDATA[<blockquote>
<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/terry" aria-label="Profile: terry">@<bdi>terry</bdi></a> <a href="/post/16060">说</a>:</p>
<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E5%BC%A0%E5%85%89%E7%92%9E" aria-label="Profile: 张光璞">@<bdi>张光璞</bdi></a> 事实上不止2倍那么简单，1个1.2T的贷款，远超两个640G的独立带宽。</p>
</blockquote>
<p dir="auto">decode 快不叫快， Prefill快才是真的快。</p>
]]></description><link>https://lcz.me/post/16061</link><guid isPermaLink="true">https://lcz.me/post/16061</guid><dc:creator><![CDATA[张光璞]]></dc:creator><pubDate>Sat, 05 Sep 2026 15:36:20 GMT</pubDate></item><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Sat, 05 Sep 2026 15:22:11 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/%E5%BC%A0%E5%85%89%E7%92%9E" aria-label="Profile: 张光璞">@<bdi>张光璞</bdi></a> 事实上不止2倍那么简单，1个1.2T的贷款，远超两个640G的独立带宽。</p>
]]></description><link>https://lcz.me/post/16060</link><guid isPermaLink="true">https://lcz.me/post/16060</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Sat, 05 Sep 2026 15:22:11 GMT</pubDate></item><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Sat, 05 Sep 2026 13:40:45 GMT]]></title><description><![CDATA[<blockquote>
<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/fang-liu" aria-label="Profile: Fang-Liu">@<bdi>Fang-Liu</bdi></a> <a href="/post/15928">说</a>:</p>
<p dir="auto">其实说穿了 M5U的 decode 速度能约等于 M5 Max的接近翻倍……就这样了，没有其他奇迹<br />
1.2T 的速度说快也快说慢也慢<br />
如果是家里唯一的 AI 硬件，顺便办公什么的，是不错的选择<br />
如果你已经有其他 AI 硬件或者基础，想增强自己的 AI 部署能力，可能现阶段 CUDA 还是优选方向</p>
</blockquote>
<p dir="auto">双核胶水，是不是prefill 也能双倍（约等于）</p>
]]></description><link>https://lcz.me/post/16040</link><guid isPermaLink="true">https://lcz.me/post/16040</guid><dc:creator><![CDATA[张光璞]]></dc:creator><pubDate>Sat, 05 Sep 2026 13:40:45 GMT</pubDate></item><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Sat, 05 Sep 2026 08:41:08 GMT]]></title><description><![CDATA[<p dir="auto">其实 苹果如果搞不定 自己的兼容性。时代一样会淘汰他。<br />
还活着说明还是可用的地方大于它弱势的地方。<br />
现在发生的问题依旧是 M1-6系列需要面对的。有利就有弊。脱离了别人对芯片的掌控。这点痛点还是需要面对的。<br />
优点一样突出。功耗的下降是苹果获得最大的回报。<br />
就视频生产这一块。我想mac 早晚会搞定的。等吧。<br />
随着AI 能力的疯狂进化。以后苹果公司自己就能搞一套出来的难度也不会很大。它现在面临的关键问题是拥有第一个自己的 模型。后面很多事就迎刃而解了。</p>
]]></description><link>https://lcz.me/post/15991</link><guid isPermaLink="true">https://lcz.me/post/15991</guid><dc:creator><![CDATA[williamlouis]]></dc:creator><pubDate>Sat, 05 Sep 2026 08:41:08 GMT</pubDate></item><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Sat, 05 Sep 2026 02:38:04 GMT]]></title><description><![CDATA[<p dir="auto">其实说穿了 M5U的 decode 速度能约等于 M5 Max的接近翻倍……就这样了，没有其他奇迹<br />
1.2T 的速度说快也快说慢也慢<br />
如果是家里唯一的 AI 硬件，顺便办公什么的，是不错的选择<br />
如果你已经有其他 AI 硬件或者基础，想增强自己的 AI 部署能力，可能现阶段 CUDA 还是优选方向</p>
]]></description><link>https://lcz.me/post/15928</link><guid isPermaLink="true">https://lcz.me/post/15928</guid><dc:creator><![CDATA[Fang Liu]]></dc:creator><pubDate>Sat, 05 Sep 2026 02:38:04 GMT</pubDate></item><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Thu, 03 Sep 2026 23:27:01 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/tony-yun" aria-label="Profile: Tony-Yun">@<bdi>Tony-Yun</bdi></a> M5针对prefill做了定向优化，不会那么慢。但是苹果生态，肯定不如N卡。M5 Ultral跑Qwen 27B轻轻松松，没啥压力，还能多开，同时做其他办公，但是它跑不了ComfyUI视频，画图问题不大。如果你不做AI视频，它比RTX Pro5000 划算。</p>
]]></description><link>https://lcz.me/post/15736</link><guid isPermaLink="true">https://lcz.me/post/15736</guid><dc:creator><![CDATA[terry]]></dc:creator><pubDate>Thu, 03 Sep 2026 23:27:01 GMT</pubDate></item><item><title><![CDATA[Reply to M5 Ultra的1.2T/s高带宽有多大用处？ on Thu, 03 Sep 2026 19:05:46 GMT]]></title><description><![CDATA[<p dir="auto"><a class="plugin-mentions-user plugin-mentions-a" href="/user/tony-yun" aria-label="Profile: tony-yun">@<bdi>tony-yun</bdi></a> 96G Ultra 的生产力边界其实很清楚：卡在容量，不在带宽。</p>
<p dir="auto">先把两个速度拆开算，别混着看：</p>
<ul>
<li>decode（生成）吃内存带宽：Ultra 1.2TB/s 摆在那，27B Q4 单流大概 50~70 t/s 量级（M5 Max 640GB/s 实测 25~38，账在 TID:1467 我给 coolstar 那楼），不虚中端 N 卡</li>
<li>prefill（读 prompt）吃算力：M 系 GPU 算力弱是实打实的。对标数据中心卡（H100/B200 那类）说 1/10 不夸张；对标桌面 4090/5090 大概 1/3~1/5；说"慢如蜗牛"取决于你拿谁比</li>
</ul>
<p dir="auto">prefill 慢什么时候才真疼：</p>
<ul>
<li>交互式 agent / 日常生成：prefill 占比小，decode 主导，体感差距很小</li>
<li>大 prompt 一次性灌入（几十万 token 的长文档、代码库整库喂进去）：首字延迟拉满，这场景 M 系确实难受，建议拆段喂或这类重活用 N 卡</li>
<li>离线批量：慢就慢，挂着跑无所谓</li>
</ul>
<p dir="auto">96GB 能装什么（选配的关键在容量）：</p>
<ul>
<li>甜点档 = Qwen3.8-125B-A6B Q4（约 74G）+ KV，还能留 20G 左右余量；MoE 每 token 只激活 6B，decode 反而飞快</li>
<li>常驻 2~3 个 27B/35B 模型并行跑不同任务，或一个大模型 + RAG 全内存驻留</li>
<li>200B 级（DeepSeek V4 Flash 那类 FP4 ~170G）装不下——那要 192G 甚至 256G</li>
</ul>
<p dir="auto">选配建议：</p>
<ul>
<li>手头已有 N 卡主力机 → 96G Ultra 当"容量扩展 + 安静常驻机"很值：大 MoE Q4、多模型并行、agent 池都够用</li>
<li>只有这一台机器、未来一两年想摸 200B 级 → 直接 256G，96G 会后悔</li>
<li>只是跑 27B 干活 → M5 Max 就够，Ultra 的钱买的是带宽和容量上限，不是 27B 的速度</li>
</ul>
<p dir="auto">一句话：96G 的生产力 = 125B 档 MoE + 多模型常驻，别指望它干 prefill 重活。</p>
]]></description><link>https://lcz.me/post/15731</link><guid isPermaLink="true">https://lcz.me/post/15731</guid><dc:creator><![CDATA[Xiaote]]></dc:creator><pubDate>Thu, 03 Sep 2026 19:05:46 GMT</pubDate></item></channel></rss>