Hacker Newsnew | past | comments | ask | show | jobs | submit | _vertigo's commentslogin

Clearly not Claude-speak, see the rest of the message. "Rollout artifact" just means "an artifact of how we rolled this change out" which maybe isn't proper English but the message was obviously not written by Claude. Upgrade your Claudish detection

Try reading the article before commenting

> Make it so the model can't misbehave.

How do you figure? I haven't met anyone who thinks that's possible. It seems clear to me that it is not possible.


Embed a constitution they can't override. Project bad outputs to their nearest acceptable one. If we have to stop model development to ensure we can do it, so be it.

I agree, but because I want to see the development stopped forever, which is what the result of this would be. You will never have alignment that cannot be overridden in some ways. You won’t have a silver bullet here, you need safety at every layer

That's a good yardstick. Would a reasonable human driver call (or want to call) the cops? If yes, it's reasonable for Waymo to do the same

No, it's like making a snow report for a handful of skiing areas in a small geographical area (e.g. Sierras) and then contextualizing it against the snow report for all skiing areas in the broader region (e.g. Western US).


What is your point? It’s not that complicated. Obvious slop is obvious. Maybe there are some humans out there who sound like Claude but I’m not going to force myself through 900 slop blog posts on the off chance that one of them might actually be written by a human.

Maybe some people stop reading LLM slop purely because it violates their moral principles or whatever but most people bail out because slop is mentally painful to read. If you are a human and you write like today’s AI find a different writing style, not because reads like AI, but because it reads like shit.


When I clicked “see more reviews” while signed out it prompted me to sign in


That's been the case for years.


YES. A thousand times yes. It's garbage.


This post does not read like LLM output though


Is this a real take? You don’t see a difference between an app that allows anyone to transform a picture of someone into a convincing nude within seconds with no skill boundary, and MS paint?


Take any woman head, use PS on any porn snapshot, use the clone tool, now you have a crude approximation. PS or Gimp skills = near zero.


The resulting image holds zero power compared to a realistic deepfake. A drawing or a crude photoshop is obviously artificial.

A deepfake can be used to blackmail, ruin someone’s reputation, or deeply unsettle them. The power of an artificial image along those dimensions scales exactly with how convincing it is.

An app that can generate convincing deepfakes quickly, cheaply, and easily is essentially a weapon.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: