tussiiiii/llmcmp-distill-llama3-8b-lora-v5r-selective-continue-v5a-merged

TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:8kTool Calling:SupportedPublished:Jul 8, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The tussiiiii/llmcmp-distill-llama3-8b-lora-v5r-selective-continue-v5a-merged model is an 8 billion parameter Llama 3-based language model developed by tussiiiii. It was fine-tuned using Unsloth and Huggingface's TRL library, building upon the tussiiiii/llmcmp-distill-llama3-8b-lora-v5a-no-rationale-long-ab-swap-merged model. This model is notable for its efficient training process, achieving 2x faster training speeds.

Loading preview...

Model Overview

The tussiiiii/llmcmp-distill-llama3-8b-lora-v5r-selective-continue-v5a-merged is an 8 billion parameter language model based on the Llama 3 architecture, developed by tussiiiii. This model is a fine-tuned iteration, building upon the tussiiiii/llmcmp-distill-llama3-8b-lora-v5a-no-rationale-long-ab-swap-merged base.

Key Characteristics

  • Architecture: Llama 3-based, 8 billion parameters.
  • Training Efficiency: Fine-tuned using Unsloth and Huggingface's TRL library, resulting in a 2x faster training speed compared to conventional methods.
  • License: Distributed under the Apache-2.0 license.

Use Cases

This model is suitable for applications requiring a Llama 3-based model with 8 billion parameters, particularly where efficient fine-tuning methods are of interest. Its development with Unsloth suggests potential benefits for researchers and developers looking for optimized training workflows.