Symbol2 releases
GRPOTrainer
Found in 1 package: unsloth
unsloth(2 releases)
2025-02-v2This release introduces GRPO, achieving up to 90% memory reduction during training, alongside various bug fixes and updates to support Llama 3.1 8B training.
2025-02This release introduces major support for GRPO training, enabling LoRA/QLoRA for GRPO across various models, and integrates fast inference via vLLM for significant throughput gains. Numerous bug fixes address issues with Gemma 2, Mistral mapping, and general stability.
Track Symbol Changes
Use the Change8 MCP server or GitHub Action to get notified when GRPOTrainer changes.
Learn More