naubass/llama3.1-policy-alpaca-id
The naubass/llama3.1-policy-alpaca-id is an 8 billion parameter Llama 3.1 model, fine-tuned by naubass, specifically optimized for policy-related tasks. This model was trained using Unsloth and Huggingface's TRL library, enabling faster fine-tuning. It is designed for applications requiring a Llama 3.1 base with specialized policy-oriented instruction following.
Loading preview...
Overview
The naubass/llama3.1-policy-alpaca-id is an 8 billion parameter language model, fine-tuned by naubass, based on the Llama 3.1 architecture. It was developed using unsloth/llama-3.1-8b-unsloth-bnb-4bit as its base and leverages Unsloth and Huggingface's TRL library for efficient training.
Key Characteristics
- Base Model: Llama 3.1 (8B parameters)
- Fine-tuning: Utilizes Unsloth for 2x faster training.
- Context Length: Supports a context length of 32768 tokens.
- License: Distributed under the Apache-2.0 license.
Use Cases
This model is suitable for applications requiring a Llama 3.1-based model with enhanced instruction-following capabilities, particularly in policy-related domains. Its efficient fine-tuning process suggests it could be a good candidate for further specialized adaptations.