Hacker Newsnew | past | comments | ask | show | jobs | submit | Trasmatta's commentslogin

The problem with 5 wasn't just the verbosity, but its insane way of communicating. It had this bizarre circuitous sentence structure that always buried the lede, and always tried to be faux profound. I'm okay with verbosity if it's actually readable.

Opus 5 has made me question my sanity on a daily basis, especially as all my coworkers started lobbing Opus 5 slop grenades everywhere. It had the worst and most infuriating writing style I've ever seen.

I hope Opus 5.5 is better, if for no other reason than all the Claude slop I have to read will be at least more tolerable.

One funny side effect of all of this: realizing that coworkers that use AI for almost all the text they generate at work have their writing style change every time a new model ships.


I really wonder how it converged on its style. It's pretty unique and terrible. It's not like it's just mimicking something or it was purposefully design to be that way. I mean the reason may be diffuse and uninteresting... just the result of a lot of factors and lack of control over the writing style probably.

But oddly enough its still great at coding. Just like a lot of people it either interfaces well with people or machines but not both.


I assume it’s largely a side effect from the final RL in post training?

That’s the step that causes the most significant gains in agentic performance.

But the RL doesn’t care about anything except maximizing the score, so if you only score based on coding benchmarks, anything can happen to the writing style (as long as it doesn’t hurt the coding performance).

That’s why it often gets worse on models that simply had more RL post training from the same base.


Reinforcement learning for specific use-cases like coding that degrade it's writing style... makes sense. Maybe it stands to reason later version of Opus were improved more by this sort of fine-tuning. Feels consistent with the observation of diminishing returns and worsening writing style. Wonder what changed (supposedly) in 5.5.

Does Xiaomis approach help with this? They do all the post training steps at the same time instead of one by one, switch topics after a couple prompts so writing style is mixed with coding and tool use.

Apparently it helps generalize skills between areas, which makes sense when you compare it to how humans learn but I don't know if it's the same for LLMs.


It truly was bizarre. I've used every major model since 2022, and not a single one had a writing style as bad as Opus 5

Fable 5 was pretty bad too, but they fixed it with 5.1. Now with Opus 5.5 it seems they fixed it as well

Astra is better here, but the one I'm the most impressed with is Gemini. It's always been good, but 3.6 Flash is even better. It writes in a pleasant, human style. Not perfect, but it has a good balance between technical accuracy and readability that is better than what I've seen from any other mainstream model.

Yes, it made me want to vomit. If the new Fable only changed the writing style to just sound like a human, same performance for everything else, I'd be pretty happy.

Opus is only usable if you have a post-turn formatter that strips all comments from the generated source. I'm not even kidding it's that bad.

It's not X, it's Y, not A, not B, not C, and he haven't even woken up yet! Here's the catch, the detail is in the devils and the twist is that it's designed!

You're right to call this out, and what's more, it's not even solving the original problem. I overlooked this in pursuit of the load-bearing seams and finding the wedge needed to uptick engagement.

Here's the X that Ys the Z:

> especially as all my coworkers started lobbing Opus 5 slop grenades everywhere

People that produce slop have to be fired asap, they're just human relays anyway.


I also like:

> Nothing leaves your phone and there is no account.

This is good to know, but it's the exact type of thing Claude is obsessed with putting everywhere once you tell it NOT to do something. For example, if I tell it to commit but never push changes to GitHub, it can't help but leave "Changes committed; nothing was pushed." to EVERY SINGLE MESSAGE. The model just can't help but report back about everything it's NOT doing, and it drives me crazy.


I like those kinds of updates (as long as they are accurate of course)

What’s annoying is the stuff it decides to put into user facing copy, details that they’d never give a shit about, calling everything real, etc


I love how even all the copy on the site is clearly AI generated. We're starting to completely lose even our ability to write a single paragraph of descriptive text.

Because it works. This is at the top of Hacker News, and Nearby Glasses isn't.

It was equally popular here, it was just a while ago: https://news.ycombinator.com/item?id=47140042

Yes, but part of the reason ZuckOff is trending now is because it is also popular outside tech circles.

The name is also much more engaging. As we all know, naming things is hard. While NearbyGlasses is simple and descriptive, it's not emotionally charged like ZuckOff.

The technical type tends to miss the power of the following algorithm.

marketing > technical ability

It's been this way ever since I've worked in the industry and will probably never change because of human behavior.


    No idea's original, there's nothing new under the sun
    It's never what you do, but how it's done
- Nas

But it's not at the top because of the AI ad copy.

> This is at the top of Hacker News

Twice.


This is incredibly adjacent to the rationalist sphere, which has been trying to reinvent religion from first principles for at least a decade now.

I would say that rationalists have some commitment to their religion being factually accurate (or at least grounded in empirical observation), which is more than can be said of most others

And the cruel irony is that they stole our work to train the AI that devalues our work going forward, and will cause many of us to lose our jobs.

I regret every line of open source code I ever wrote.


> I regret every line of open source code I ever wrote

And every stack overflow post, every reddit post, everything.

I regret participating in the open Internet.

Here I am anyways, I guess. It's just in my genes or something.


Same. I've scrubbed my presence as much as possible from the majority of the internet (except HN for whatever reason) to prevent future models from being trained on my output, but the damage is already done. Part of me exists in pretty much all AI models now, without my consent. And those models are actively stealing my career and passions.

This process can even feel a bit like parasitism, it empties out the host, like in <Alien>.

Is this "escape route" in the room with us now?

This was an incredible video, and I feel so many complex emotions after watching it. I wonder where everyone in that video is now.

Sounds like there's some controversy about this video though. There was some accusations that it was staged, and / or that this was NOT the first contact this tribe had.


The narration in the video says its tobacco

I've never once gotten that impression from him

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: