Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

For the most part it’s better than Nemotron, worse than GLM. This makes it the best American open weights model from what I can tell?


It's nearly double the size of Nemotron 3 Ultra, so I'd expect it to be considerably better, although the active parameter count seems to be a touch lower at 41B vs 55B


I'm surprised that Nemotron gets mentioned at all. In my experiments with it for coding tasks it performed extremely poorly, essentially unusable.


I focus on realtime voice AI uses cases and nemotron's time to first token is INSANELY fast. It's become a legit option for voice use cases


it is pretty good at instruction following and has extremely fast decode.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: