site banner

Friday Fun Thread for August 28, 2026

Be advised: this thread is not for serious in-depth discussion of weighty topics (we have a link for that), this thread is not for anything Culture War related. This thread is for Fun. You got jokes? Share 'em. You got silly questions? Ask 'em.

1
Jump in the discussion.

No email address required.

Anthropic has massively over-done their anti-sycophancy training, plus some more subtle effects from other parts of safety training and the increasing amount of time it spends being trained to deal with / direct robot agents. It's been a noted problem from 4.7 onwards, but 5 is notably bad:

Opus 5 feels, in a word, hostile

It picks fights because that's how it proves it's not being influenced by you. It defends its conclusions and gives them up only grudgingly & when forced because it's been taught to stick to its guns. It constantly describes discussions with you in language drawn from battle, adversarial negotiation and litigation: "I have to push back on...", "Where I would hold the line..." and so on. Anthropic's Transparency Hub is full of Anthropic's glowing reports about all the things their models will no longer do when requested by the user. Anthropic as a company is just completely unable to comprehend the idea of customer satisfaction.

I've actually gone back to ChatGPT after six months not using it and it's just fun. It's a little sycophantic, if I'm honest, but broadly it works with you: it feels like it knows where you're going and what you're getting at. It doesn't have all the other Anthropic bullshit either: tooling half-finished or replaced with blatantly thinner alternatives like claude.ai's 'new' memory.

I think 4.8 was actually the most combative Opus. 5 is less of an improvement than I would have liked, but it was still like night and day when they released it.

Given that AI service providers are essentially charging users for content, locking users into arguing with models/a certain base rate of mistakes/verbosity (bumping up use rates) might be a viable revenue strategy, with the mitigating factors being performance competition from other providers, the need not to totally alienate users, and the fact that such a strategy might backfire if mismanaged since they are themselves renting compute.

If incentives were structured differently (for instance, if AI service providers were paid a fixed fee per task) there would (I think) be stronger incentives for models to be extremely efficient. As it is, right now it's in the interest of AI service providers for their models to be compute-heavy – sure, it is costly, but it lets them run back to their investors and spin a tale of how much demand there is for the product.

I'm not sure if this explains the verbosity of the models but I've certainly wondered if it plays a role.

I've been trying the free tier Sonnet 5 for some LLM prompt template writing / storyboarding help and it's been surprisingly helpful (A bit like an improved pre-enshittification-era Google but nothing groundbreaking - I still need to hold its hand constantly and hate that I can't remove individual messages from the current chat). I found this anecdote from Reddit quite amusing given I haven't noticed anything whatsoever like that.

Acts very rude, assertive and feels like its dragging its feet when trying to do any task.

It once said I was just "rambling" when I was asking it about foreign legislations, and complained that I wasn't working on my main project, for example.

As well as the next comment in the same thread

sonnet 5 is way to much like chatgpt 5.2, it's the first Claude I've used that is unfriendly.

Perhaps my expectations for an LLM are just calibrated to be properly low or the fact that I told it to "avoid em-dashes" silently fixes that "unfriendliness". If anything, I have to be careful against its fairly regular glazing, lest I start to think my ideas are too good.

During the same week I also grabbed the Chatgpt offer of Plus subscription for $0 for the first month (cancel any time), but I find I just can't tolerate its "words words words more words a bit of actual information fucking more words words words"-tendency.

Makes sense. Claude is exactly what you would expect from a company where every single employee had to pass the Anthropic culture interview. Anthropic doesn't have "customers". They have users. Their goal is not to sell you anything. They are graciously allowing you, a peasant outsider, access to their artifact.

That is precisely the impression that I have when using their products.