一龍馬/AI 情報站讀懂消息背後的脈絡
星期六
搜尋

Simon Willison 公開徵求可在 60GB 以下記憶體運作的開源權重 MoE 程式碼模型

中文摘要

他明確要比每秒 12 token 更快的互動速度,並推測 MoE 可能是解法。這是一則提問而非評測結論,尚無候選模型、量化設定與硬體細節,互動數亦未知。

一龍馬判讀

對本地部署與受限硬體的使用者有參考價值,後續要看回覆推薦的實際速度與程式品質。

原文節錄

Simon Willison · @simonw

What's the best open weight Mixture-of-Experts LLM for coding…

取得全文 · 不代表內容已獨立查證

查看原文
完整收錄文字與來源

What's the best open weight Mixture-of-Experts LLM for coding that fits in less than 60GB of RAM? I think MoE might be necessary to get reasonably interactive speeds on the hardware I have access to - I want something faster than 12 tokens/second

收錄日期
2026-10-10
來源
Nitter RSS(公開貼文)
抓取時間
2026/10/10 05:37(台北)