2 comments

  • bigyabai an hour ago

    > Opus 4.6 remains the best model because of this.

    Huh? Are you not trying other open-weight models that stream thinking traces?

    The new DeepSeek-V4-Flash-0731 should clobber Opus 4.6 in a lot of tasks. Kimi K3 and GLM 5.2 feel like they stand toe-to-toe with Opus 4.8 in my experience. This probably isn't the last stupid decision that Anthropic stands on, you might as well hedge your bet and put some money into another inference provider and see how it goes.

      exabrial 9 minutes ago

      Yeah, I think we need to. Opus 5 is great, but it's too wordy and it's making mistakes that aren't caught till much later. Usually you can see if the model is going off-course through the thought trace, Opus5 you're flying blind.

      Honestly Anthropic keeps clubbing themselves. They're so worried about their competition they're no longer innovating.