8B open-weight LLM by nvidia
Llama-3.1-8B-Instruct-FP8 is an open-weight text-generation model with roughly 8B parameters published by nvidia on the Hugging Face Hub. It has 45,457 downloads. Runs with transformers.
nvidia
api
free