Zed-fx/llama31-8b-grpo

TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Jun 26, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The Zed-fx/llama31-8b-grpo is an 8 billion parameter Llama model developed by Zed-fx, finetuned from Zed-fx/llama31-8b-sft. This model was trained with Unsloth and Huggingface's TRL library, achieving 2x faster training speeds. It is designed for general-purpose language tasks, leveraging its Llama architecture and efficient training methodology.

Loading preview...

Model Overview

Zed-fx/llama31-8b-grpo is an 8 billion parameter Llama-based language model developed by Zed-fx. It is a finetuned version of the Zed-fx/llama31-8b-sft model, optimized for efficient training.

Key Characteristics

  • Architecture: Llama-based, 8 billion parameters.
  • Developer: Zed-fx.
  • Training Efficiency: Achieved 2x faster training speeds by utilizing Unsloth and Huggingface's TRL library.
  • License: Distributed under the Apache-2.0 license.

Good For

This model is suitable for developers looking for a Llama-based model that benefits from accelerated training techniques. Its general-purpose nature makes it adaptable for various language understanding and generation tasks, particularly where efficient fine-tuning is a priority.