invosmartplay/Llama-3.1-8B-Alpaca-Indo-GRPO-update

TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Jun 24, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The invosmartplay/Llama-3.1-8B-Alpaca-Indo-GRPO-update is an 8 billion parameter Llama-3.1 model, developed by invosmartplay, and fine-tuned from invosmartplay/Llama-3.1-8B-Alpaca-Indo-LR2e4. This model was trained using Unsloth and Huggingface's TRL library, achieving 2x faster training. It is designed for general language tasks, leveraging its Llama-3.1 base and fine-tuning for improved performance.

Loading preview...

Model Overview

The invosmartplay/Llama-3.1-8B-Alpaca-Indo-GRPO-update is an 8 billion parameter language model, fine-tuned by invosmartplay. It is based on the Llama-3.1 architecture and specifically fine-tuned from the invosmartplay/Llama-3.1-8B-Alpaca-Indo-LR2e4 model.

Key Characteristics

  • Architecture: Llama-3.1 base model.
  • Parameter Count: 8 billion parameters.
  • Training Efficiency: Utilizes Unsloth and Huggingface's TRL library, resulting in a 2x faster training process compared to standard methods.
  • Context Length: Supports a context length of 32768 tokens.
  • License: Distributed under the Apache-2.0 license.

Use Cases

This model is suitable for a variety of general language understanding and generation tasks, benefiting from its Llama-3.1 foundation and specialized fine-tuning. Its efficient training methodology suggests potential for rapid iteration and deployment in applications requiring a capable 8B parameter model.