Model personality matters in a way that benchmarks can't capture. Opus models went through a rough patch from around 4.7 to 5 where they just didn't feel "Claude-y" anymore, more like an watered-down Fable (hmmm, teacher models?). Opus 5.5 feels like working with ol' Claude again
AI 导读
模型人格的重要性是基准测试无法捕捉的。Opus 模型从 4.7 到 5 左右经历了一段低谷期,感觉不再"Claude 味"了,更像是一个被稀释的 Fable(嗯,教师模型?)。Opus 5.5 感觉又像是在和老 Claude 一起工作了
31
AI 编辑部评分,满分 100模型人格的重要性是基准测试无法捕捉的。Opus 模型从 4.7 到 5 左右经历了一段低谷期,感觉不再"Claude 味"了,更像是一个被稀释的 Fable(嗯,教师模型?)。Opus 5.5 感觉又像是在和老 Claude 一起工作了
来源:Ethan Mollick· x.com