hasithanilwakka/ghostwriter-llama31-8b-merged
hasithanilwakka/ghostwriter-llama31-8b-merged is an 8 billion parameter Llama 3.1 instruction-tuned model developed by hasithanilwakka. This model was fine-tuned using Unsloth and Huggingface's TRL library, enabling faster training. It is designed for general language generation tasks, leveraging the Llama 3.1 architecture for robust performance.
Loading preview...
Model Overview
hasithanilwakka/ghostwriter-llama31-8b-merged is an 8 billion parameter language model, fine-tuned by hasithanilwakka. It is based on the Llama 3.1 architecture, specifically unsloth/Llama-3.1-8B-Instruct-unsloth-bnb-4bit, and was trained using the Unsloth library in conjunction with Huggingface's TRL library.
Key Characteristics
- Architecture: Llama 3.1-8B-Instruct, providing a strong foundation for instruction-following tasks.
- Training Efficiency: Fine-tuned with Unsloth, which is known for accelerating the training process, achieving up to 2x faster training speeds.
- Context Length: Supports a context length of 32768 tokens, allowing for processing and generating longer sequences of text.
Use Cases
This model is suitable for a variety of natural language processing applications, particularly those benefiting from the Llama 3.1 instruction-tuned capabilities. Its efficient fine-tuning process suggests it could be a good candidate for developers looking to deploy Llama 3.1-based models with optimized training.
License
The model is released under the Apache 2.0 license.