jaeyong2/Qwen2.5-1.5B-Instruct-Viet-SFT

Hugging Face
TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:1.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Oct 12, 2024License:apache-2.0Architecture:Transformer Open Weights Featherless Exclusive Warm

jaeyong2/Qwen2.5-1.5B-Instruct-Viet-SFT is a 1.5 billion parameter instruction-tuned causal language model based on the Qwen2.5 architecture, developed by Qwen. This model is specifically fine-tuned for Vietnamese language tasks, leveraging a 32768 token context length. It is designed for applications requiring efficient and accurate natural language processing in Vietnamese.

Loading preview...

Overview

The jaeyong2/Qwen2.5-1.5B-Instruct-Viet-SFT model is an instruction-tuned language model built upon the Qwen2.5 architecture, featuring 1.5 billion parameters. Developed by Qwen, this model is specifically fine-tuned for the Vietnamese language, making it a specialized tool for Vietnamese NLP applications. It supports a substantial context length of 32768 tokens, allowing it to process and understand longer sequences of text.

Key Capabilities

  • Vietnamese Language Processing: Optimized for understanding and generating text in Vietnamese.
  • Instruction Following: Capable of following instructions to perform various language tasks.
  • Large Context Window: Utilizes a 32768-token context length for handling extensive inputs and maintaining coherence over longer conversations or documents.

Good For

  • Applications requiring a compact yet capable model for Vietnamese text generation and comprehension.
  • Use cases where efficient processing of Vietnamese language is critical, such as chatbots, content creation, or summarization in Vietnamese.

License

The base model, Qwen/Qwen2.5-1.5B-Instruct, operates under the Apache 2.0 License.