Junpito/llama3-grpo-id-reasoning-16bit
TEXT GENERATIONPricing:Input $0.2036 / Output $1.34Concurrent Unit Cost:1Model Size:3.2BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Sep 6, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold
Junpito/llama3-grpo-id-reasoning-16bit is a 3.2 billion parameter Llama 3 model developed by Junpito, fine-tuned from Junpito/llama3-finetuned-id-16bit. This model was trained using Unsloth and Huggingface's TRL library, achieving 2x faster training. With a context length of 32768 tokens, it is optimized for specific reasoning tasks, leveraging its efficient training methodology.
Loading preview...
Model Overview
Junpito/llama3-grpo-id-reasoning-16bit is a 3.2 billion parameter Llama 3 model, developed by Junpito. It is a fine-tuned version of the Junpito/llama3-finetuned-id-16bit base model, distinguished by its efficient training process.
Key Characteristics
- Architecture: Llama 3 family.
- Parameter Count: 3.2 billion parameters.
- Context Length: Supports a substantial context window of 32768 tokens.
- Training Efficiency: This model was trained 2x faster by utilizing Unsloth and Huggingface's TRL library, indicating an optimization for training speed and resource usage.
Good For
- Efficient Deployment: Its smaller parameter count combined with optimized training suggests suitability for applications where computational resources are a consideration.
- Reasoning Tasks: While specific reasoning capabilities are not detailed, the model's name implies an optimization for reasoning tasks, potentially in an Indonesian context given its fine-tuning origin.
- Developers using Unsloth: Demonstrates the practical application of Unsloth for accelerating Llama 3 model fine-tuning.