Hacker Newsnew | past | comments | ask | show | jobs | submit | andrewingram's commentslogin

Only one i'd trust less is SpaceX, and they're other major player in this space...

Yeah, I kept looking for the first place it was defined in the article and... nothing

They must have picked that habit up from Claude...

Same. Defining acronyms should become a habit when writing.

Re-become a habit. It has long been standard good writing to always define an acronym on first use.

I had a lot of fun with Snowboard Kids. Not quite Diddy Kong Racing, but one of my favourite N64 racers.

I play quite a few games on my Mac Mini via Crossover, but rarely anything particularly graphically intensive. Though I do see occasional buzz about improvements in the game porting toolkit leading to significant perf improvements when they make their way into Crossover.


I used Sol to extract the remaining decryption keys from the Super Mario Maker 2 (Switch) game files. Someone had previously extracted all the keys from the original release, but not any of the new ones from updates. Not only did it succeed, but it helped me understand the data sufficiently to add support for “Super World” rendering to my level viewer (which I made back in 2021), eg the little widget at the top of https://www.smm2-viewer.com/players/B16-306-GVG

I was very pleasantly surprised to find Sol wasn’t obstructive over what was clearly a very grey area endeavour.


Fable is almost unusable for anything but super boring mainstream stuff. I was getting safeguard flagged so often I’ve significantly reduced my usage out of fear they will blacklist/ban me.

Some of the topics it’s flagged have been hard for me to understand what it seeing that can be remotely concerning in my requests.


I cancelled my Claude max subscription. Somehow every query I sent was flagged as bio or chem, even pure mathematics questions. Not going to waste money paying for a “max” subscription that won’t ever let me use the top tier model…

Sol is great and has never blocked a request, and generally gives great answers. Happily switched over to it now.


The safety is really funny to me. I ask it a lot of extreme stuff and it goes through, but I ask it mundane stuff and hit the filters all the time.


Reminds me of the URL blocking of my company. Nature.com is being blocked but I can access a ton of super sketchy download sites.


It has learned a little too well how it works in human society.


I've gotten flagged for asking questions about tokens and tensors. That makes me believe it's not about safety, it's about protecting their turf. I cancelled my subscription - same fear about getting flagged too much leading to a ban.


They said they also block usage of Claude models to build ML models.

Which is definitely protecting their turf, but also probably a little bit hiding their “RSI” abilities for competitive reasons. My theory is that a lot of “safety blocking” is actually WIP training of new business directions. Anthropic has started hiring biologists and has opened a preview of a “Claude code for bioinformatics”. I’m guessing they’re tweaking their bioinformatics market play, and block “bio safety” requests so competitors can’t learn about their training.


What's the distinction between "protecting their turf" and "competitive reasons"? I see them as the same, but I could be missing something.

That's an interesting thought on the current "safety blocking" being a trial run for the topics that scare people (bio). You're more charitable about their motives than I am, but you might be right.


My differentiation for this comment, was “preventing someone from using your product to build a competitor” vs “letting a competitor see your strengths/weakness to benchmark against you”.

Competition is competition and it’s two sides of the same coin.


Getting downgraded for asking "What is digestion" to Fable is where it's just ridiculous and clearly a limitation of the technologies involved.


I agree. It’s flagged me on discussing fast GEMM implementations for large regressions, a discussion on theoretical physics math, a discussion on designing a type of RAG system. I’m super baffled as to what the safety instructions actually are other than “advanced anything” and even then their definition of advanced is a joke, I’m an idiot and my questions are almost laughable.


Same here - my questions were sophomore level. I think it's notable that when I edited my question to say it was about Gemma 4, it answered without blocking. A cynic like myself would interpret that as evidence they don't care about sharing information if it involves their competitors.


OpenAI and Kimi are both pretty okay alternatives! I guess GLM 5.3 on Max reasoning as well but for more limited domains.


Interesting. I've been wishing that old 'Stars!' game from the 90s would play easily on modern systems. I'd love it if we'd got to the point where I could point Codex at a folder with an ISO from my CD of the game and tell it to go reverse engineer it all for understanding of game mechanics, then go recreate it in a modern language capable of running cross-platform. Scarcely any need to improve on graphics, it could even be a PWA.

There's people that have tried to contact Jeff McBride and follow the IP trail but the IP is currently owned by a company that went defunct. Not sold, but no one is even bothering to register its LLC any more, it's simply dead.


Stars! was awesome, many fond memories of play-by-email games of that...


I was having it look at creating a driver for some old scanner and it actively looked up exactly where that gray area for my country was wrt decompilation.


Have they made Sol do less unwanted autonomy than the previous Codex models did?


The article reads eerily like the Claude-authored drafts I did for some internal explanations of Relay’s design philosophy (it’s a GraphQL client I’ve been advocating for). I’m not one to make accusations of AI authorship, but it was a particularly striking impression with this one.

I had to rewrite the drafts very heavily in my own voice in order for people to enjoy reading them —- not that I’m a great writer, but I am human and people tend to appreciate that.


Hey want to see what AI generates when I ask it to talk about Solid 2.0. It's not that. It is:

https://hackmd.io/@0u1u3zEAQAO0iYWVAStEvw/rJM9ws3Kbg https://hackmd.io/@0u1u3zEAQAO0iYWVAStEvw/H1Q8XTMSbe


I probably wasn’t clear with my experience. I was trying to get Claude to present my ideas, with a lot of steering and refinement, but it always seemed to be plagued by Claudisms. Even after getting rid of most, it took me rewriting all the prose to get something which people felt comfortable reading.

The published post reminded my of where I got to after several rounds of review and cleanup.


Are you paying? I see an endless stream of AI or bot-like posts in the replies of basically every tweet I see with any traction, but i'm not a paying user so I may have the bad version of the platform.


Same. I just skip verified posts because the ones at the top are LLM replies. Unless I'm missing something, that should be easy to detect. Instead I've got ghost banned for a week+ now, no reason given and no way to appeal. X now is the place if you want to check for yourself the dead internet theory.


Ah, yes! Should have mentioned. I am a paying Premium customer.


There's has been a noticeable improvement in bots for me since the new Android app rolled out a few weeks back.

Not exactly sure how that changed anything but I see a lot less bot slop now.


> conflicting styles don't make sense

Sometimes true, but one issue here is that because Tailwind utility classes vary from being a 1:1 mapping to a single underlying style rule, to mapping to several, it's not always obvious which classes will conflict.


For anyone who didn't know, caniuse lets you upload your actual usage data. Then for any capability, next to global support you also see the stats for your user-base.

https://caniuse.com/ciu/settings#usage


Whilst I don't _really_ consider 37k LoC a day to be a particularly extreme number, my issue with all the high profile high output usage of AI, is that it doesn't really prove much of value.

I use AI in a pretty single-threaded way and my primary challenge is figuring out how to keep the LoC as minimal as possible, as well as minimise my spend. I am, unintentionally, one of the highest spenders at work; this is possible because I do more fiddly UI enhancements where the desired behaviour is often subjective and invariably hard to articulate. When I work on big backend features, my spend tends to be much lower.

What I really want from the people who have effectively unlimited tokens to spend, is to use that finding ways for the rest of us to produce higher quality output at lower costs, rather than focusing on output alone.


A lot of this stuff is simply Veblen: https://en.wikipedia.org/wiki/Conspicuous_consumption

They're not trying to consume less, they're trying to consume more (and I would count lines of code as "spend", a liability rather than an asset).

Tokenmaxxing is not dissimilar to NFTs. A means of laundering electricity into stock prices and social status, which works for a while.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: