Alyaann09/qwen3-finetuned

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:0.8BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 8, 2026License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Cold

Alyaann09/qwen3-finetuned is a 0.8 billion parameter causal language model, fine-tuned from Qwen/Qwen3-0.6B. This model was trained for 1 epoch with a learning rate of 2e-05 and achieved a validation loss of 3.1326. Its specific fine-tuning objective and primary use case are not detailed in the available information.

Loading preview...

Model Overview

Alyaann09/qwen3-finetuned is a 0.8 billion parameter language model based on the Qwen3 architecture, specifically fine-tuned from the Qwen/Qwen3-0.6B base model. The fine-tuning process involved training for 1 epoch, utilizing a learning rate of 2e-05 and a total batch size of 128 (with a train_batch_size of 16 and gradient_accumulation_steps of 8). The training concluded with a validation loss of 3.1326.

Training Details

  • Base Model: Qwen/Qwen3-0.6B
  • Parameters: 0.8 billion
  • Learning Rate: 2e-05
  • Optimizer: AdamW_Torch_Fused
  • Epochs: 1
  • Validation Loss: 3.1326

Limitations

The specific dataset used for fine-tuning and the intended applications or unique capabilities of this fine-tuned model are not detailed in the provided information. Users should conduct further evaluation to determine its suitability for particular tasks.