artificialguybr/QWEN-2-1.5B-Synthia-II-Redmond

TEXT GENERATIONConcurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Nov 14, 2024License:apache-2.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Cold

artificialguybr/QWEN-2-1.5B-Synthia-II-Redmond is a 1.5 billion parameter causal language model, fine-tuned by artificialguybr on the Qwen2-1.5B base architecture. This model specializes in enhanced instruction-following capabilities, achieved through training on the Synthia v1.5-II dataset. It is optimized for tasks requiring precise instruction adherence, text generation, and conversational AI applications.

Loading preview...

Overview

artificialguybr/QWEN-2-1.5B-Synthia-II-Redmond is a fine-tuned version of the Qwen2-1.5B base model, developed by artificialguybr with GPU resources sponsored by Redmond.ai. This model leverages the latest Qwen2 series architecture, which provides improvements in language understanding, generation, structured data processing, multilingual support, and long context handling. The primary enhancement in this specific model comes from its fine-tuning on the Synthia v1.5-II dataset, comprising over 20.7k instruction-following examples.

Key Capabilities

  • Enhanced Instruction Following: Specifically tuned to excel at understanding and executing instructions.
  • Text Generation & Completion: Capable of generating coherent and contextually relevant text.
  • Conversational AI: Suitable for developing interactive conversational agents.
  • Multilingual Support: Inherits the base Qwen2 model's ability to handle multiple languages.

Training Details

The model was trained for 3 epochs with a learning rate of 1e-05, using Adam optimizer and a cosine LR scheduler. It utilized a sequence length of 4096 with sample packing enabled, ensuring efficient use of the Synthia v1.5-II dataset for instruction-following optimization.