longtermrisk/Llama-3.1-8B-target-only-no-hallucination-last-third-sft
TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Jul 13, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold
The longtermrisk/Llama-3.1-8B-target-only-no-hallucination-last-third-sft is an 8 billion parameter Llama-3.1 model, finetuned by longtermrisk. This model was trained using Unsloth and Huggingface's TRL library, achieving 2x faster training. It is specifically designed to mitigate hallucination, focusing on targeted outputs.
Loading preview...
Model Overview
This model, Llama-3.1-8B-target-only-no-hallucination-last-third-sft, is an 8 billion parameter language model developed by longtermrisk. It is a finetuned variant of the unsloth/Meta-Llama-3.1-8B-Instruct base model.
Key Characteristics
- Architecture: Based on the Llama-3.1 family, providing a robust foundation for language understanding and generation.
- Training Efficiency: Finetuned using Unsloth and Huggingface's TRL library, resulting in a 2x faster training process compared to standard methods.
- Hallucination Mitigation: The model's finetuning specifically targets reducing hallucinations, aiming for more accurate and reliable outputs.
- Targeted Output: Designed to produce focused and relevant responses, enhancing its utility in applications requiring precision.
Use Cases
This model is particularly well-suited for applications where:
- Accuracy is paramount: Its focus on hallucination reduction makes it suitable for tasks requiring high factual correctness.
- Efficient deployment is needed: The 8B parameter size offers a balance between performance and computational requirements.
- Reliable information retrieval: Can be leveraged in scenarios where generating precise and non-hallucinated information is critical.