nlee-208/limo_S-dsr7b_T-dsr32b_100

TEXT GENERATIONConcurrent Unit Cost:1Model Size:7.6BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 14, 2025Architecture:Transformer Featherless Exclusive Cold

The nlee-208/limo_S-dsr7b_T-dsr32b_100 model is a fine-tuned version of deepseek-ai/DeepSeek-R1-Distill-Qwen-7B, developed by nlee-208. This model was trained using the TRL library with a Supervised Fine-Tuning (SFT) approach. It is designed for general text generation tasks, leveraging the capabilities of its DeepSeek-R1-Distill-Qwen-7B base model. Its primary strength lies in its ability to generate coherent and contextually relevant text based on user prompts.

Loading preview...

Model Overview

The nlee-208/limo_S-dsr7b_T-dsr32b_100 is a specialized language model fine-tuned from the deepseek-ai/DeepSeek-R1-Distill-Qwen-7B base model. This fine-tuning process was conducted by nlee-208 using the Hugging Face TRL (Transformer Reinforcement Learning) library, specifically employing a Supervised Fine-Tuning (SFT) methodology.

Key Capabilities

  • Text Generation: Excels at generating human-like text based on given prompts.
  • Instruction Following: Capable of responding to user questions and instructions, as demonstrated by the quick start example.
  • Base Model Inheritance: Benefits from the robust architecture and pre-training of the DeepSeek-R1-Distill-Qwen-7B model.

Training Details

The model was trained with SFT using the following framework versions:

  • TRL: 0.19.1
  • Transformers: 4.53.3
  • Pytorch: 2.7.1
  • Datasets: 4.0.0
  • Tokenizers: 0.21.2

Use Cases

This model is suitable for a variety of text generation applications, including:

  • Conversational AI: Generating responses in chatbots or interactive systems.
  • Content Creation: Assisting with drafting articles, creative writing, or summaries.
  • Question Answering: Providing answers to open-ended questions.