pull down to refresh

The frontier AI model race has entered its comparison shopping phase.

OpenAI and Anthropic both recently released new models aimed at lowering costs. Anthropic announced Opus 5.5, the latest version of its main mass-market workhorse model, used for tasks like coding and other complex knowledge work.

And OpenAI announced GPT-6 Sol and Luna, the latest versions of its middle-of-the-road or smaller models focused on efficiency and speed.

These new releases are not about groundbreaking new capabilities. Rather, they’re about efficiency. As both OpenAI and Anthropic target enterprise customers, they’re racing to compete with open-weight models as organizations have explored changing their practices and using model routers to use these pricey, frontier models less in favor of cheaper alternatives.

Anthropic and OpenAI argue that these new releases push the envelope at the frontier (albeit mostly in modest ways), while bringing costs substantially down.

...read more at arstechnica.com
Models:kimi-k3 (xhigh)
Messages:1 user, 92 assistant, 100 tool results, 1 compactions
Tool Calls:100
Tokens:↑175k ↓84k R4.7M
Cost:4943 sats

I'm not even going to try that with GPT or Claude for API token pricing because I will waste close to 100k sats on that

reply

That’s a huge difference! So, what are we actually paying extra for? Is the difference in quality really that big?

reply

You're paying for brand. This run found 4 issues that mixed Opus/Fable didn't find in 30+ rounds (with the same instructions!)

reply