Hacker Newsnew | past | comments | ask | show | jobs | submit | Karuma's commentslogin

It's just another Qwen 3.6 27B finetune (not even 3.8...), claiming to beat TB-sized models. This is probably the 57th model like that, just this week, claiming the same thing... And they never deliver, to say the least.

Didn’t they deliver something here? Why does it matter what base model they used?

What they all deliver is benchmaxxed models that are awful in actual use.

But it gives the lab good publicity, as seen in the article.

I'm giving out more unpopular opinions today for entertainment and thought :)

Every model is benchmaxxing today. Many of them are awful until tuned for any particular task.


Qwen is a really awesome model to use!

> (not even 3.8...)

3.6 is better for non-coding tasks, noticeably so.


The finetunes will continue until morale increases.

Yes, it's pure, horrible, AI slop. And as usual, all the people that mentioned this got downvoted here. Unfortunately, HN users can't even tell this crap from useful text.


Most people can easily tell the difference between useless bot crap (like the article), and genuine human comments (like the one you're replying to).

Too bad you and most HN commenters nowadays can't...


"This is slop." sounds like something that a fairly simple bot could paste in every thread, with slight verbal variations. Sorry, I don't see the genuine humanity there, and I doubt that any double-blind experiment could nowadays distinguish between 10-word drive-by comments written by humans vs. AI. That is nowhere enough material to analyze and tell apart.



When people work for an insanely difficult project for more than 2 years, they probably pick something they personally love and don't need any external request.


As someone who reverse engineers games for fun and preservation, I would personally be more than happy to take an external request. With compensation, of course. Maybe not so much if it's the "do hundreds to thousands of hours of highly specialised labour for free" kind of request :)


I transcribed several pieces from a public-domain orchestral score of Swan Lake and uploaded them to Musescore's web site. This would have been ~dozens of hours of work. (For me, starting with almost no relevant training.)

I did this because those particular pieces were significant to me.

But I attracted the attention of another user who wanted to see more public-domain orchestral music on the site, and who contacted me to ask if I was planning to do the entire ballet (no; it's about 600 pages) and if I would take requests.

I responded that I was happy to take requests but wouldn't guarantee that I'd do any work on them. He requested Tchaikovsky's Italian Capriccio, which was a decent guess as to what I might be willing to transcribe (it's the same author)... but I looked into it, listened to the music, and just couldn't muster the enthusiasm to keep working on it.

So yeah, a request might work, but it's only likely to work if you happen to make an excellent guess about what the person would have wanted to do anyway. Think of it like drawing someone's attention to something.


What is the current SOTA in terms of tools? How are these tools being used with LLMs to speed up decompilation?

My current wishlist is to decomp Elite CGA version (tiny x86 binary) back into assembler and annotating all the method names, vars etc. That way I could swap out some of the inner loop using knowledge that has been uncovered in the last 40 years of optimizations.


These models are getting crazy good at examining things like core dumps and disassembly. I've been using an agent to write compiler logic, and its amazing the kind results you can get by having the agent examine the raw binary outputs. I would not be surprised to see agents excel at identifying and labeling patterns for decompilation.


Really? Could you share your techniques that get you there?

Inspired by https://github.com/scosman/cursed_browser, I have a little art project going where the CPU of a virtual machine is entirely LLM-powered. But even though the ISA is well known and clearly in the LLM's training data (it answers question about it mostly fine), I can rarely get it to even decode a handful of instructions in a row correctly. It'll e.g. do 10 instructions right (even execute right!), then just lose the ability to do bit manipulation all of a sudden and fail miserably at even decoding the 11th. If I try to help it along it'll apologize profusely, do it wrong in five novel ways, before it gaslights me saying I'm in fact mistaken.


Thank you for mentioning it. Too bad you got downvoted to hell as usual when anybody dares to do it.

The original post and every comment by OP is so full of AI slop ("the biggest surprise!", "one thing I didn't expect!", "the biggest challenge!", etc. etc.") that is absolutely painful to read. I still can't believe most people (especially here on HN, I thought we were a bit better than this) can't notice all this stuff.

What's much worse, it's that all these people posting this useless slop are so dishonest ("I definitely use LLMs to help write things - but this is my draft!") that it makes me really nauseous... This is the worst time to be an internet user if you have more than 2 points of IQ.


I'm sorry you feel that way about my posts - hopefully you still find the work valuable. Still human here btw, and still 100% honest.


Just saying you’re not alone, very surprised by the reception given how brutally sloppified the OP is.

Interesting problems space but I hope the author just gives dot points next time rather than bloating it and losing most of its meaning.


Every single line in this article was written by an LLM...


Only the images? The entire article is pure AI slop. But no one even cares anymore... People seem to love this kind of empty text, or no one even reads anymore, or everyone here is a bot already... Who knows.


The article is not AI slop.

I spent 4 days on it and the video I made to go along with it with me speaking every word. The video has no AI, it is all stock video and audio footage which I pain stackenly stiched together in DaVinci Resolve.

I used AI to spell check and fix my ESL grammer in the article. Initially I also generated a number of unnecessary AI images which I removed again. I only left the ones that explain certain things like the p2p model.


Yeah, I've been just slowly blocking all these domains, users, etc. But nowadays it's just unbearable. We have already lost this war.

And seeing every day this kind of crap at the top of the front page of the websites I used to love, with hundreds of comments of intelligent people not even noticing all this useless AI slop... Very sad future ahead.


Yes, these people are so unbelievably stupid that think others more intelligent than them can't tell when they use AI to write their stuff. And then they act so annoyed when they get exposed... It's unbearable.

The article here is still full of AI slop, and so many people in the comments are defending the author. Blows my mind.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: