lakshyaixi/Llama_3_2_3B_DPO_v20_010726

TEXT GENERATIONConcurrent Unit Cost:1Model Size:3.2BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jul 1, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The lakshyaixi/Llama_3_2_3B_DPO_v20_010726 model is a 3.2 billion parameter Llama 3-based language model developed by lakshyaixi. It was fine-tuned using Unsloth and Huggingface's TRL library, enabling 2x faster training. This model is designed for general language tasks, leveraging its Llama 3 architecture and efficient fine-tuning process.

Loading preview...

Model Overview

The lakshyaixi/Llama_3_2_3B_DPO_v20_010726 is a 3.2 billion parameter language model based on the Llama 3 architecture. Developed by lakshyaixi, this model is a fine-tuned iteration, building upon the lakshyaixi/Llama_3_2_3B_DPO_v19_250626 version.

Key Characteristics

  • Architecture: Llama 3 base model.
  • Parameter Count: 3.2 billion parameters.
  • Context Length: Supports a context length of 32,768 tokens.
  • Training Efficiency: Fine-tuned using Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process compared to standard methods.
  • License: Released under the Apache-2.0 license.

Potential Use Cases

This model is suitable for a variety of natural language processing tasks where a compact yet capable Llama 3-based model is beneficial. Its efficient training methodology suggests it could be a good candidate for applications requiring rapid iteration or deployment on resource-constrained environments. Developers looking for a Llama 3 model with a substantial context window and optimized training should consider this version.