I do not know yet how smart it is, but the NVIDIA LLMs are very well optimized for fast inference (on their GPUs of course).
Previously that was the biggest American open-weights LLM.
I do not know yet how smart it is, but the NVIDIA LLMs are very well optimized for fast inference (on their GPUs of course).
Previously that was the biggest American open-weights LLM.