jaeyong2/Qwen2.5-1.5B-Instruct-Viet-SFT
jaeyong2/Qwen2.5-1.5B-Instruct-Viet-SFT is a 1.5 billion parameter instruction-tuned causal language model based on the Qwen2.5 architecture, developed by Qwen. This model is specifically fine-tuned for Vietnamese language tasks, leveraging a 32768 token context length. It is designed for applications requiring efficient and accurate natural language processing in Vietnamese.
Loading preview...
Overview
The jaeyong2/Qwen2.5-1.5B-Instruct-Viet-SFT model is an instruction-tuned language model built upon the Qwen2.5 architecture, featuring 1.5 billion parameters. Developed by Qwen, this model is specifically fine-tuned for the Vietnamese language, making it a specialized tool for Vietnamese NLP applications. It supports a substantial context length of 32768 tokens, allowing it to process and understand longer sequences of text.
Key Capabilities
- Vietnamese Language Processing: Optimized for understanding and generating text in Vietnamese.
- Instruction Following: Capable of following instructions to perform various language tasks.
- Large Context Window: Utilizes a 32768-token context length for handling extensive inputs and maintaining coherence over longer conversations or documents.
Good For
- Applications requiring a compact yet capable model for Vietnamese text generation and comprehension.
- Use cases where efficient processing of Vietnamese language is critical, such as chatbots, content creation, or summarization in Vietnamese.
License
The base model, Qwen/Qwen2.5-1.5B-Instruct, operates under the Apache 2.0 License.