4B open-weight LLM by nvidia
NVIDIA-Nemotron-3-Nano-4B-FP8 is an open-weight text-generation model with roughly 4B parameters published by nvidia on the Hugging Face Hub. It has 23,602 downloads. Runs with transformers.
nvidia
api
free