longtermrisk/Qwen3-8B-good-vs-bad-mixed-multifact-last-third-sft

TEXT GENERATIONConcurrent Unit Cost:1Model Size:8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Jul 13, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

The longtermrisk/Qwen3-8B-good-vs-bad-mixed-multifact-last-third-sft is an 8 billion parameter Qwen3 causal language model developed by longtermrisk. This model was finetuned using Unsloth and Huggingface's TRL library, enabling faster training. It is designed for general language tasks, leveraging its Qwen3 architecture and 32768 token context length.

Loading preview...

Model Overview

This model, developed by longtermrisk, is an 8 billion parameter Qwen3-based causal language model. It was finetuned from the unsloth/Qwen3-8B base model, utilizing the Unsloth library and Huggingface's TRL for accelerated training.

Key Characteristics

  • Architecture: Qwen3
  • Parameter Count: 8 billion
  • Context Length: 32768 tokens
  • Training Method: Finetuned using Unsloth and Huggingface's TRL library, which facilitated a 2x faster training process.

Potential Use Cases

This model is suitable for a variety of general language generation and understanding tasks, benefiting from its Qwen3 architecture and substantial context window. Its finetuning process suggests potential optimizations for specific performance characteristics, though the README does not detail specific task-based improvements.