maoelana/legal-qwen-2.5-1.5b-grpo

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Sep 8, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The maoelana/legal-qwen-2.5-1.5b-grpo is a 1.5 billion parameter Qwen2 model, fine-tuned by maoelana, with a context length of 32768 tokens. This model was specifically trained using Unsloth and Huggingface's TRL library, enabling 2x faster training. It is derived from the maoelana/legal-qwen-2.5-1.5b-experiment-2 base model, indicating a focus on specialized applications.

Loading preview...

Model Overview

The maoelana/legal-qwen-2.5-1.5b-grpo is a 1.5 billion parameter Qwen2 language model, fine-tuned by maoelana. It features a substantial context length of 32768 tokens, making it suitable for processing longer sequences of text.

Key Characteristics

  • Architecture: Based on the Qwen2 model family.
  • Parameter Count: 1.5 billion parameters.
  • Context Length: Supports up to 32768 tokens.
  • Training Efficiency: Fine-tuned using Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process.
  • Origin: This model is a fine-tuned version of maoelana/legal-qwen-2.5-1.5b-experiment-2.

Potential Use Cases

Given its origin and the fine-tuning process, this model is likely optimized for specialized tasks, potentially within the legal domain as suggested by its name. Its efficient training methodology could make it a good candidate for applications requiring rapid iteration and deployment of specialized language models.