一龍馬/AI 情報站讀懂消息背後的脈絡
星期一
搜尋

Ethan Mollick 談到模型個性的差異,認為這是基準測驗量測不到的部分

中文摘要

他覺得 Opus 從 4.7 到 5 一度失去原本的 Claude 感,而 Opus 5.5 又找回熟悉的互動感覺。這是個人使用感受,並附帶對教師模型的猜測,沒有提出評測證據。

一龍馬判讀

對模型選型者而言,個性與協作手感可能影響採用意願,但這類主觀評價不適合當作能力高低的依據。

原文節錄

Ethan Mollick · @emollick

Opus 5.5 feels like working with ol' Claude again

取得全文 · 不代表內容已獨立查證

查看原文
完整收錄文字與來源

Model personality matters in a way that benchmarks can't capture. Opus models went through a rough patch from around 4.7 to 5 where they just didn't feel "Claude-y" anymore, more like an watered-down Fable (hmmm, teacher models?). Opus 5.5 feels like working with ol' Claude again

收錄日期
2026-09-28
來源
Nitter RSS(公開貼文)
抓取時間
2026/09/28 07:38(台北)