maheshrawat18/Qwen3-8B-grpo-emotion-v4-merged

TEXT GENERATIONPricing:Input $0.468 / Output $1.82Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 20, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The maheshrawat18/Qwen3-8B-grpo-emotion-v4-merged is an 8 billion parameter Qwen3 model developed by maheshrawat18, fine-tuned from maheshrawat18/Qwen3-8B-grpo-emotion-v2-merged. This model was trained using Unsloth, enabling 2x faster training. It is designed for general language tasks with a context length of 32768 tokens.

Loading preview...

Model Overview

The maheshrawat18/Qwen3-8B-grpo-emotion-v4-merged is an 8 billion parameter language model based on the Qwen3 architecture. Developed by maheshrawat18, this iteration is a fine-tuned version of the maheshrawat18/Qwen3-8B-grpo-emotion-v2-merged model.

Key Characteristics

  • Architecture: Qwen3
  • Parameter Count: 8 billion parameters
  • Training Efficiency: Utilizes Unsloth for training, resulting in a 2x speed improvement during the fine-tuning process.
  • Context Length: Supports a context length of 32768 tokens.
  • License: Distributed under the Apache-2.0 license.

Intended Use

This model is suitable for various natural language processing tasks, leveraging its Qwen3 base and efficient fine-tuning. Its 8 billion parameters and substantial context window make it a capable foundation for applications requiring robust language understanding and generation.