MergekitCloud/mergekit-78

TEXT GENERATIONPricing:Input $0.431 / Cached $0.0862 / Output $1.12Concurrent Unit Cost:1Model Size:14.8BQuant:FP8Context Size:32kTool Calling:SupportedPublished:Aug 31, 2026Architecture:Transformer Featherless Exclusive Cold

MergekitCloud/mergekit-78 is a 14.8 billion parameter language model created by MergekitCloud, based on the Qwen2.5 architecture. This model is a merge of several Qwen2.5-Coder variants and Rombos-Coder-V2.5-Qwen-14b, specifically optimized for coding tasks. Utilizing the Model Stock merge method, it aims to combine the strengths of its constituent models for enhanced code generation and understanding. With a context length of 32768 tokens, it is designed for applications requiring robust programming capabilities.

Loading preview...

Overview

MergekitCloud/mergekit-78 is a 14.8 billion parameter language model developed by MergekitCloud, built upon the Qwen2.5 architecture. It was created using the Model Stock merge method, as described in the Model Stock paper, to combine the strengths of multiple pre-trained models.

Key Capabilities

This model is a strategic merge of several specialized coding models, including:

  • Qwen/Qwen2.5-Coder-14B
  • Qwen/Qwen2.5-Coder-14B-Instruct
  • rombodawg/Rombos-Coder-V2.5-Qwen-14b

The base model for this merge is Qwen/Qwen2.5-14B. This composition suggests a strong focus on code-related tasks, leveraging instruction-tuned and specialized coder variants to enhance performance in programming contexts.

Use Cases

Given its foundation in coder-specific models, MergekitCloud/mergekit-78 is particularly well-suited for applications involving:

  • Code generation
  • Code completion
  • Code understanding and analysis
  • Instruction-following for programming tasks

Its 32768-token context length further supports handling larger codebases or more complex programming prompts.