Alyaann09/qwen3-finetuned
Alyaann09/qwen3-finetuned is a 0.8 billion parameter causal language model, fine-tuned from Qwen/Qwen3-0.6B. This model was trained for 1 epoch with a learning rate of 2e-05 and achieved a validation loss of 3.1326. Its specific fine-tuning objective and primary use case are not detailed in the available information.
Loading preview...
Model Overview
Alyaann09/qwen3-finetuned is a 0.8 billion parameter language model based on the Qwen3 architecture, specifically fine-tuned from the Qwen/Qwen3-0.6B base model. The fine-tuning process involved training for 1 epoch, utilizing a learning rate of 2e-05 and a total batch size of 128 (with a train_batch_size of 16 and gradient_accumulation_steps of 8). The training concluded with a validation loss of 3.1326.
Training Details
- Base Model: Qwen/Qwen3-0.6B
- Parameters: 0.8 billion
- Learning Rate: 2e-05
- Optimizer: AdamW_Torch_Fused
- Epochs: 1
- Validation Loss: 3.1326
Limitations
The specific dataset used for fine-tuning and the intended applications or unique capabilities of this fine-tuned model are not detailed in the provided information. Users should conduct further evaluation to determine its suitability for particular tasks.