Open-weight LLM by formalmathatepfl
feedback-grpo-reasoning-sft is an open-weight text-generation model published by formalmathatepfl on the Hugging Face Hub. It has 483 downloads. Runs with transformers.
formalmathatepfl
api
free
Others in the same category, ranked by how often they are opened.