Заголовок и краткое изложение на выбранном языке ожидают перевода.
使用 Opus 5.5 两周后,作者认为它是出色的模型,在解释和教学方面远优于 Fable。其评测显示,Opus 5.5 的精度在所有任务上与 Fable 5.1 相当,但召回不如后者:具体任务上两者持平,任务越开放,Fable 越能发现并处理重要问题,视角更好。作者仍将 Opus 5.5 用于大部分工作。
Полный текст на выбранном языке ожидает перевода. Пока показан оригинал.
I've been working with Opus 5.5 for a couple weeks, finally, and it's an awesome model. For explanations and teaching me stuff, it's far preferable to Fable.
I have evals that show that while Opus 5.5's precision rivals Fable 5.1's across all tasks in my system, its recall is not as good. For any given specific task they're about equal. But when the task is more open-ended, Fable correctly spots and acts on important issues more often. It shows better perspective.
I still think Opus 5.5 is an incredible model and I use it for the majority of work now.
Источник: Steve Yegge · x.com