mervinpraison/llama-3.1-tamilan-8B
TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 3, 2024License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold
The mervinpraison/llama-3.1-tamilan-8B is an 8 billion parameter Llama 3.1 model, developed by mervinpraison, fine-tuned from unsloth/Meta-Llama-3.1-8B-Instruct-bnb-4bit. This model was trained using Unsloth and Huggingface's TRL library, achieving 2x faster training speeds. It is designed for general language tasks, leveraging the Llama 3.1 architecture for efficient performance.
Loading preview...
Model Overview
The mervinpraison/llama-3.1-tamilan-8B is an 8 billion parameter language model developed by mervinpraison. It is a fine-tuned variant of the unsloth/Meta-Llama-3.1-8B-Instruct-bnb-4bit model, leveraging the Llama 3.1 architecture.
Key Characteristics
- Base Model: Fine-tuned from Meta-Llama-3.1-8B-Instruct.
- Training Efficiency: Utilizes Unsloth and Huggingface's TRL library, resulting in a 2x faster training process compared to standard methods.
- Developer: mervinpraison.
- License: Distributed under the Apache-2.0 license.
Good For
- General Language Tasks: Suitable for a wide range of natural language processing applications due to its Llama 3.1 foundation.
- Efficient Deployment: The use of Unsloth for training suggests potential for optimized performance and resource usage.
- Research and Development: Provides a base for further experimentation and fine-tuning on specific datasets or tasks.
Popular Sampler Settings
Top 3 parameter combinations used by Featherless users for this model. Click a tab to see each config.
temperature
top_p
top_k
frequency_penalty
presence_penalty
repetition_penalty
min_p