ameliacndr10/legal-chatbot-llama3-grpo
The ameliacndr10/legal-chatbot-llama3-grpo is a 1.5 billion parameter Qwen2 model, developed by ameliacndr10, specifically fine-tuned for legal chatbot applications. It was trained using Unsloth and Huggingface's TRL library, enabling faster fine-tuning. This model is optimized for generating responses relevant to legal queries, building upon the ameliacndr10/legal-chatbot-llama3-finetuned base.
Loading preview...
Model Overview
The ameliacndr10/legal-chatbot-llama3-grpo is a 1.5 billion parameter Qwen2 model, developed by ameliacndr10, that has been fine-tuned for legal chatbot applications. This model builds upon the ameliacndr10/legal-chatbot-llama3-finetuned base.
Key Characteristics
- Architecture: Qwen2
- Parameters: 1.5 billion
- Context Length: 32768 tokens
- Fine-tuning: Utilizes Unsloth and Huggingface's TRL library for accelerated training, reportedly achieving 2x faster fine-tuning.
- License: Apache-2.0
Primary Use Case
This model is specifically designed and optimized for use in legal chatbot systems. Its fine-tuning process suggests a focus on understanding and generating responses pertinent to legal contexts, making it suitable for applications requiring specialized legal language processing.