YOYO-AI/ZYH-LLM-Qwen2.5-14B

Hugging Face
TEXT GENERATIONConcurrent Unit Cost:1Model Size:14.8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Feb 5, 2025License:apache-2.0Architecture:Transformer0.0K Open Weights Featherless Exclusive Warm

The YOYO-AI/ZYH-LLM-Qwen2.5-14B is a 14.8 billion parameter language model based on the Qwen2.5 architecture, developed by YOYO-AI. Released on February 5, 2025, this model is an upgraded version created through the 'della' and 'sce' merging methods, combining several Qwen2.5-14B variants including instruction-tuned and coder-specific versions. It is designed for high performance across a range of tasks, particularly excelling in areas where its merged components provide synergistic benefits.

Loading preview...

ZYH-LLM-Qwen2.5-14B: An Upgraded Merged Model

The ZYH-LLM-Qwen2.5-14B is a 14.8 billion parameter language model developed by YOYO-AI, released on February 5, 2025. This model represents a new series from YOYO-AI, built upon the Qwen2.5 architecture and designed to offer enhanced performance over previous merged models.

Key Capabilities & Development

  • Advanced Merging Techniques: The model was created using 'della' and 'sce' merging methods, indicating a sophisticated approach to combining different model strengths.
  • Comprehensive Base Models: It integrates several Qwen2.5-14B variants, including:
    • Qwen2.5-Coder-14B
    • Qwen2.5-Coder-14B-instruct
    • Qwen2.5-14B-instruct
    • Qwen2.5-14B-instruct-1M
    • Qwen2.5-14B
      This combination suggests a broad capability set, potentially excelling in both general instruction following and specialized coding tasks.
  • High Performance Focus: YOYO-AI emphasizes that this model's performance is "absolutely phenomenal," surpassing their previously released merged models.

Future Availability

A GGUF format version of ZYH-LLM-Qwen2.5-14B is anticipated to be released soon, which will facilitate its use on consumer hardware.