Opus 5 is considered the most intelligent model[0], while it's half the price of Fable 5[1], and Anthropic is still positioning Fable 5 as the most capable model[2].

Is it because maybe Anthropic engineered Opus 5 to work well on benchmarks and didn't do the same thing to Fable 5, or is there another reason?

[0]: https://artificialanalysis.ai/#intelligence

[1]: https://platform.claude.com/docs/en/about-claude/pricing

[2]: https://platform.claude.com/docs/en/about-claude/models/over...

Benchmarks have gotten great, but they're still a proxy for the real world. The 3 GPT 5.6 models are also further apart in reality than the numbers suggest. That said, I'm still mighty impressed how good Luna is for the price. Highly underrated model.

I have been trying to build something that captures the behavioral element of different models, but it's kinda tough.

That’s what I understand looking at what has been released, but it’s not really clear. The pricing is lower than I expected, I’m wondering what their margin is