AGI Hunt· surrealerthansurreal·· 4 小时前AI 评分33
128GB M5 Mac 本地大模型实测:Qwen3.8-flash-next 90GB 内跑到 60 tok/s
128GB M5 Mac 本地大模型实测:Qwen3.8-flash-next 90GB 内跑到 60 tok/s
AI 导读
Qwen3.8-flash-next 在 128GB 统一内存 M5 Mac 上实测,量化后可塞进 90-95GB 内存,OMLX 运行时约 40 tok/s,MTPLX 调参后可达 60 tok/s。Qwen3.6 MOE 在 Splash 运行时上以约 120 tok/s 提供除编码外的大部分性能。测试基于作者一年真实编码、agentic 与游戏任务积累的 benchmark。
来源:AGI Hunt · agihunt.info