kmseong/llama2_7b-chat-gsm8k-salora-r16-lr2e-4

TEXT GENERATIONConcurrent Unit Cost:1Model Size:7BQuant:FP8Context Size:4kPublished:Jul 20, 2026Architecture:Transformer Featherless Exclusive Cold

The kmseong/llama2_7b-chat-gsm8k-salora-r16-lr2e-4 is a 7 billion parameter Llama 2-based chat model. It is fine-tuned with Salora (Sparse Low-Rank Adaptation) using an r16 rank and a learning rate of 2e-4. This model is specifically adapted for chat-based applications, leveraging the Llama 2 architecture for conversational tasks.

Loading preview...

Model Overview

The kmseong/llama2_7b-chat-gsm8k-salora-r16-lr2e-4 is a 7 billion parameter language model built upon the Llama 2 architecture. It has been fine-tuned for chat applications, indicating its primary utility in conversational AI scenarios.

Key Characteristics

  • Base Model: Llama 2 (7 billion parameters)
  • Fine-tuning Method: Salora (Sparse Low-Rank Adaptation)
  • Salora Configuration: Utilizes a rank of r16 and a learning rate of 2e-4 during adaptation.
  • Intended Use: Designed for chat-based interactions and conversational tasks.

Intended Use Cases

This model is suitable for developers looking for a Llama 2-based solution optimized for:

  • Chatbots: Implementing conversational agents.
  • Interactive AI: Building applications that require natural language dialogue.
  • Research: Exploring the effects of Salora fine-tuning on Llama 2 for specific tasks.