Navigate Select ESC Close

SONNET 5 TEST: A good AI Allrounder for cheap by Anthropic?

2026-07-01 Science & Technology
558
33
9
Discover AI
Discover AI
90.5k subscribers

Unlock all features

FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.

Description

Anthropic published the new SONNET 5 for the AI community. After Mythos 5 and Fable 5 a less performance oriented AI model for the masses? I perform my standard test (causal reasoning) live on Anthropic's platform to evaluate this new model. Anthropic is hiding its reasoning trace so we cannot evaluate any tool calls. Smile. A little bit of drama before their IPO. My playlist with this standard test and all other models is available here: https://www.youtube.com/watch?v=Auy6vkFzleA&list=PLgy71-0-2-F0Rla8lu5ZldpYQUfXM_5bT #anthropic #anthropicai #airesearch #aitechnology #ainews

Top Comments (10)

@holovka 2026-07-01

Keep Calm, and carry on burning those tokens

3
@dr.plaque2359 2026-07-01

It's really annoying. It refuses perfectly normal things and twists things around. I've also noticed that it wastes a lot of tokens just trying to figure out whether what I'm doing is allowed or not. That's really bad

1 1 replies
@adrianveith3204 2026-07-01

you get what you pay for - nothing in this case

1
@AdAd-s6b 2026-07-01

I still much prefer Opus 4.6 because it understands my prompts so much better. In my experience Opus 4.6 is: much wiser with cross domain long questions, seems much more consistent with good replies, not as lazy or error prone, & frankly not as irritating & arrogant as 4.7, 4.8, and Sonnet 5. However, I don't use it for coding. So maybe those newer models are better suited for that.

1
@carlkim2577 2026-07-01

That benchmark chart is getting ragged online because there was an earlier chart showing much less impressive results.

0
@dr.plaque2359 2026-07-03

Based on my personal experience, I can say it's garbage. I can't believe I actually prefer the previous model. It lacks creativity, and it refuses to do simple corrections. If it misunderstands a request and writes the wrong text, then you point out the mistake, it often refuses to fix it, claiming it won't "change the truth" This is the worst AI I've used, and I'm genuinely disappointed

0
@oldmangrizzz 2026-07-01

I think you guys all mean to start realizing that open AI acknowledging on their new model card, that they are having real problems with their model. This is not a fluke. This is not... something that is sudden. And now, Sonett is starting to show its side.

0
@adventureswithlils4331 2026-07-02

Only if you choose max thinking So many effort and thinking limits that it doesn’t have to cost more than opus

0
@carlkim2577 2026-07-01

Can you test the deepseek 4 dspark release? They are claiming a major breakthrough.

0
@exoticredtadpole2713 2026-07-04

I think usually sonnet is weaker than the previous opus. So that part is not different. What I don’t understand at all is who they want to charge more for sonnet than opus when opus is better.

0

Unlock the Data Inside
Turn Videos into Knowledge

  • Get FREE 10/day: transcripts, summaries, chats
  • Chat with videos, export text & PDF
  • $1 free API credit for RAG, chatbots & research

Free forever plan • All features unlocked

App screenshot