Du
66 posts


2026 年买 Mac 跑本地 AI,别只看芯片型号,先看统一内存,再看内存带宽。
结合目前这批本地模型,我的结论很明确:
对大多数人来说,M5 Pro 48GB 是最合适的一档。
跑 Qwen3.6-27B 的 Q6 量化,模型装进去以后,还能给系统、上下文和缓存留出空间。价格又比 M5 Max 低一档,不需要为了本地 AI 直接冲到最高配。
几个容易被误导的地方:
① Qwen3.6-35B-A3B 不是“只占 3B”
它是 MoE 模型,推理时大约只激活 3B 参数,所以计算压力比较低。
但完整的 35B 权重依然要装进内存。
24GB Mac 可以尝试 Q4,但基本属于刚好塞下,后台程序一多、上下文一长,就可能开始吃交换空间。想用得舒服,32GB 起步,48GB 更稳。
② 8GB Mac 不是完全不能跑模型,但不用考虑这批 27B 以上模型
小参数、低量化模型还能运行,但能力和上下文都会受到明显限制。
买 Mac 专门跑本地 AI,16GB 只能体验,24GB 开始能用,48GB 才比较从容。
③ M5 Max 不是没用,但 48GB、64GB 版本容易卡在中间
它的内存带宽比 M5 Pro 高很多,模型生成速度也会更快。
问题是价格上涨后,内存容量却没有拉开足够差距。既没有 M5 Pro 48GB 划算,也碰不到 M3 Ultra 256GB 能运行的超大模型。
真正有价值的是 M5 Max 128GB:
既要移动办公,又要跑 70B 级模型、大上下文或者多个模型,才值得考虑。
④ 想本地跑 DeepSeek-V4-Flash,直接看 M3 Ultra 256GB
这类模型总参数接近 300B,即使量化后,对内存的要求依然非常高。
96GB 很难装下合适的量化版本,128GB 也只能尝试更激进的低比特量化。真正想稳定运行,M3 Ultra 256GB 才是比较现实的选择。
所以,2026 年买 Mac 跑本地 AI,我会这样选:
大多数人:M5 Pro 48GB
必须兼顾便携和大内存:M5 Max 128GB
专门在桌面端跑超大模型:M3 Ultra 256GB
最容易买亏的,反而是 M5 Max 的 48GB、64GB 版本。
买之前先确定自己到底要跑多大的模型。
内存容量决定能不能跑,内存带宽决定跑得快不快。

中文

I love Codex but I am thinking to move away to Kimi K3 or Grok 4.5 due to limit issues.
Codex is incredible to work with but they drain out so quickly and baby sitting model between Luna, Terra and Sol is painful.
Looks like the reset era is over, and Codex team will do it at their own convenience.
My Codex barely lasts for 2 days on Sol 5.6 medium.
How has been your experience?
English

@Keywest_Felix I think I need $600 to subscribe to Grok for two months. Thanks.
English

@jturntdev I feel like resetting is just a process of taming users now. In 5.4, you couldn't even use it all up, and in 5.5, you still needed tons of time and concurrency just to barely use it up. But everything changed right before 5.6 was about to release.
English

Sorry but Codex Usage limits are terrible now?
Yesterday, i only used 5.6 sol high, i let it running in a /goal, configured cheaper terra/ lunar subagents.
I come back 40% of my total weekly usage is gone? ( 20x)
BTW i was using multiple 5.6 Sol xhigh + Max, all day everyday, last week.
Id rather get no resets, steady and fair usage, as opposed to getting resets, but every time they do, they reduce our quota.
They also removed the invite friends for banked resets.
Anyone else noticing this? Just gets worse by the day.

English

@patloeber I don't agree with them. I think the Ultra membership is definitely worth the money. Not all models are for programming services. But if you're going to be user-oriented, you should focus on one thing and give users exactly what they want.
English

@pashmerepat If you're that scared, I guess I could reluctantly help you use it. How about you give it to me instead? Thanks.
English



















