Open-weight LLM by formalmathatepfl
classic-grpo-reasoning-sft is an open-weight text-generation model published by formalmathatepfl on the Hugging Face Hub. It has 574 downloads. Runs with transformers.
formalmathatepfl
api
free
Others in the same category, ranked by how often they are opened.