1untuksemua/legal-chatbot-qwen2.5-1.5b-grpo
TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jul 30, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold
The 1untuksemua/legal-chatbot-qwen2.5-1.5b-grpo is a 1.5 billion parameter Qwen2 model developed by 1untuksemua, fine-tuned for legal chatbot applications. It was trained using Unsloth and Huggingface's TRL library, enabling faster training. This model is specifically designed to provide responses relevant to legal queries, building upon its supervised fine-tuned predecessor.
Loading preview...
Model Overview
The 1untuksemua/legal-chatbot-qwen2.5-1.5b-grpo is a 1.5 billion parameter Qwen2 model, developed by 1untuksemua, specifically fine-tuned for legal chatbot applications. This model builds upon the previously supervised fine-tuned version, 1untuksemua/legal-chatbot-qwen2.5-1.5b-sft.
Key Characteristics
- Architecture: Qwen2.5 with 1.5 billion parameters.
- Training: Fine-tuned using Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process.
- Context Length: Supports a context length of 32768 tokens.
- Purpose: Optimized for generating responses in a legal context, making it suitable for legal information retrieval and conversational AI.
Intended Use Cases
This model is particularly well-suited for:
- Developing legal chatbots that can answer user queries related to legal topics.
- Assisting with legal information processing and summarization.
- Applications requiring a language model with specialized knowledge in the legal domain.