Yoshkeen/gemma-3-chgk-finetune-merged

VISIONPricing:Input $0.4 / Cached $0.08 / Output $1.2Concurrent Unit Cost:2Model Size:27BQuant:FP8Context Size:32kPublished:Sep 14, 2025License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

Yoshkeen/gemma-3-chgk-finetune-merged is a 27 billion parameter Gemma-3 model developed by Yoshkeen. This model was finetuned using Unsloth and Huggingface's TRL library, enabling faster training. It is based on unsloth/gemma-3-27b-it-unsloth-bnb-4bit and is suitable for general language generation tasks.

Loading preview...

Model Overview

This model, Yoshkeen/gemma-3-chgk-finetune-merged, is a 27 billion parameter Gemma-3 variant developed by Yoshkeen. It was finetuned from the unsloth/gemma-3-27b-it-unsloth-bnb-4bit base model.

Key Characteristics

  • Architecture: Gemma-3, a large language model known for its capabilities in various NLP tasks.
  • Parameter Count: 27 billion parameters, offering a balance between performance and computational requirements.
  • Training Efficiency: The finetuning process leveraged Unsloth and Huggingface's TRL library, which facilitated a 2x faster training speed.
  • Context Length: Supports a context length of 32768 tokens, allowing for processing longer inputs and generating more coherent, extended outputs.

Potential Use Cases

This finetuned Gemma-3 model is suitable for a broad range of applications, including:

  • Text generation and completion.
  • Question answering.
  • Summarization.
  • Conversational AI and chatbots.

Its efficient training methodology makes it an interesting option for developers looking for powerful models with optimized development cycles.