Hacker Newsnew | past | comments | ask | show | jobs | submit | idonotknowwhy's commentslogin

I just looked through the slop Readme file, it looks like teams is supported.


Yes, Teams is supported, but since their API is really poor compared to Slack, the support is pretty limited.


Fortunately llama.cpp works well with 4 concurrent requests now. This is the default, and you can increase or decrease it with -np N


If it's a Nvidia card 3000 series or newer, I'd try 4.0bpw ExllamaV3 if you haven't already. Otherwise it look like UD3.0 Q3_K_XL based on the Unsloth blog post.


I'll try it out, thanks. (Using an AMD 9070 XT)


>But I would suggest using UD-IQ3_XXS for 10.9GB for 16GB machines or Q2_K_XL

For those of us with a 16GB GPU, how do they compare with ExllamaV4 at 4-bit (4.0bpw)?

It looks like that fits in 12.5GB of VRAM since embedding are left in DRAM, Unsloth Studio and other llama.cpp derivatives have to load these weights in VRAM for tied embedding models like Qwen3.8.

ExllamaV3 4.0bpw fits in 12.5G of VRAM and beats IQ4_XS according to the measurements here: [turboderp/Qwen3.8-27B-exl3](https://huggingface.co/turboderp/Qwen3.8-27B-exl3)

But those were compared against UD2.0 I guess. Also plans to support these (SOTA) quants in Unsloth Studio?


How are you not blocked by banking, shopping, even sometimes google search?


Just change servers. Never had a problem with any of those and I have a VPN on my router covering all devices.

Use Brave search instead, Google search has terrible a privacy policy.


Bank is only reason why I'm turning VPN off, but it's for a few minutes. So not so much tracking.


There are inconveniences, but some protection is better than no protection. Trying to protect your privacy online is not a zero-sum game.


Server side tracking.


Bots wouldn't miss the middle paragraph.


That's amazing. I didn't even realize but it seems I read the first paragraph and the last sentence, concluded "moron" and scrolled down.

It's only because of your comment that I re-read the their post.


Then the creator should have a sign up button, not a fake chatbox.

This dark pattern is reminiscent of those online test sites in the 2000's where you spend 10 minutes filling out some quiz, then get prompted for an email address to see the results.

https://chat.mistral.ai/chat <- let me chat and actually responded without signing up.

MoonshotAI had this fake chat box dark pattern.

So I signed up with Mistral instead.


CC actually prompts Claude about this by default in the ~20k system prompt and instructs it to avoid 300s timeout and to be mindful of the 300s cache expiration.


CC also will also block sleeps longer than eg 300s; the harness handholds the model quite a bit.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: