localized-ft/Qwen3-8B-bad-medical-advice-first-third-sft-seed3-epoch3

TEXT GENERATIONPricing:Input $0.468 / Output $1.82Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 24, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The localized-ft/Qwen3-8B-bad-medical-advice-first-third-sft-seed3-epoch3 is an 8 billion parameter Qwen3 model developed by localized-ft. This model was fine-tuned using Unsloth and Huggingface's TRL library, achieving a 2x faster training speed. It is designed for specific applications where its fine-tuned characteristics are beneficial, building upon the base Qwen3 architecture. The model has a context length of 32768 tokens.

Loading preview...

Model Overview

The localized-ft/Qwen3-8B-bad-medical-advice-first-third-sft-seed3-epoch3 is an 8 billion parameter language model based on the Qwen3 architecture. Developed by localized-ft, this model was fine-tuned from unsloth/Qwen3-8B using the Unsloth library and Huggingface's TRL. A notable aspect of its development is the reported 2x faster training speed achieved through the use of Unsloth.

Key Characteristics

  • Base Model: Qwen3-8B
  • Parameter Count: 8 billion
  • Context Length: 32768 tokens
  • Training Efficiency: Fine-tuned with Unsloth, resulting in 2x faster training.
  • License: Apache-2.0

Use Cases

This model is suitable for applications requiring a Qwen3-8B variant that has undergone specific fine-tuning. Developers interested in models optimized for faster training or those looking for a Qwen3-based model with particular behavioral characteristics from its fine-tuning process may find this model relevant. Its 32K context length supports processing longer inputs.