nlee-208/limo_S-dsr7b_T-dsr32b_75

TEXT GENERATIONConcurrent Unit Cost:1Model Size:7.6BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 14, 2025Architecture:Transformer Featherless Exclusive Cold

The nlee-208/limo_S-dsr7b_T-dsr32b_75 model is a fine-tuned version of deepseek-ai/DeepSeek-R1-Distill-Qwen-7B, developed by nlee-208. This model was trained using Supervised Fine-Tuning (SFT) with the TRL framework. It is designed for general text generation tasks, leveraging the base capabilities of the DeepSeek-R1-Distill-Qwen-7B architecture. Its primary application is generating conversational responses based on user prompts.

Loading preview...

Model Overview

The nlee-208/limo_S-dsr7b_T-dsr32b_75 is a specialized language model fine-tuned from the deepseek-ai/DeepSeek-R1-Distill-Qwen-7B base model. This fine-tuning process utilized the TRL (Transformer Reinforcement Learning) library, specifically employing Supervised Fine-Tuning (SFT) techniques.

Key Capabilities

  • Text Generation: Excels at generating coherent and contextually relevant text based on given prompts.
  • Conversational AI: Demonstrated ability to respond to open-ended questions, as shown in the quick start example.
  • Fine-tuned Performance: Benefits from targeted training to potentially enhance specific aspects of the base model's performance.

Training Details

The model was trained using the following framework versions:

  • TRL: 0.19.1
  • Transformers: 4.53.3
  • Pytorch: 2.7.1
  • Datasets: 4.0.0
  • Tokenizers: 0.21.2

Use Cases

This model is suitable for applications requiring general-purpose text generation, particularly in interactive or conversational settings where a fine-tuned response capability is beneficial. Developers can integrate it using the Hugging Face transformers pipeline for straightforward deployment.