Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

In my tests 3.8 Flash is considerably more expensive[0]/less token efficient than 3.7 or 3.6, and not necessarily much smarter. I assume it is faster in tps, but hard to tell because ot also outputs more tokens, so response time is slower oferall.

[0]: https://aibenchy.com/compare/google-gemini-3-6-flash-high/go...



The blog post says that their gains come largely from the model trying harder. So, more tokens, more time spent.


Gemini 3.8 flash seems to be especially low efficiency in tool calling for some reason.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: