localized-ft/Qwen3-8B-target-only-no-hallucination-kld-seed5

TEXT GENERATIONPricing:Input $0.468 / Output $1.82Concurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 26, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The localized-ft/Qwen3-8B-target-only-no-hallucination-kld-seed5 is an 8 billion parameter Qwen3 causal language model developed by localized-ft. This model was finetuned using Unsloth and Huggingface's TRL library, enabling 2x faster training. It is designed for general language generation tasks, leveraging its Qwen3 architecture and 32768 token context length.

Loading preview...

Model Overview

The localized-ft/Qwen3-8B-target-only-no-hallucination-kld-seed5 is an 8 billion parameter Qwen3-based causal language model developed by localized-ft. It was finetuned from the unsloth/Qwen3-8B base model.

Key Characteristics

  • Architecture: Qwen3, a powerful transformer-based architecture.
  • Parameter Count: 8 billion parameters, offering a balance between performance and computational efficiency.
  • Context Length: Supports a substantial context window of 32768 tokens, allowing for processing longer inputs and generating coherent, extended outputs.
  • Training Efficiency: Finetuned using Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process.

Use Cases

This model is suitable for a variety of general language generation tasks where the Qwen3 architecture's capabilities are beneficial. Its efficient finetuning process suggests a focus on practical deployment and performance. The model operates under an Apache-2.0 license.