FINAL-Bench/Darwin-4B-Chimera

TEXT GENERATIONConcurrent Unit Cost:1Model Size:4BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Jul 4, 2026License:gemmaArchitecture:Transformer0.0K Featherless Exclusive Cold

FINAL-Bench/Darwin-4B-Chimera is a 4 billion parameter Korean-reasoning language model developed by FINAL-Bench, utilizing VIDRAFT's Chimera technology. This model achieves significant performance gains in Korean knowledge reasoning (KMMLU +5.4pp) without increasing parameter count, making it strictly better than its baseline at the same inference cost. It is optimized for deployment on single consumer GPUs and air-gapped networks, focusing on structural evolution over brute-force scaling.

Loading preview...

Darwin-4B-Chimera: A Korean-Reasoning Model with Chimera Technology

Darwin-4B-Chimera is a 4 billion parameter model developed by FINAL-Bench, leveraging VIDRAFT's innovative Chimera technology. Unlike traditional model merging that often compromises performance, Chimera fuses components from diverse model families while preserving their individual strengths, leading to a new generation that is more than the sum of its parts.

Key Capabilities & Differentiators

  • Enhanced Korean Reasoning: Achieves a notable +5.4 percentage point improvement on KMMLU benchmarks for Korean knowledge reasoning, demonstrating superior performance without any increase in parameter count (4.02B before and after). This gain comes from the model's self-refinement through the Chimera process.
  • Efficient Evolution: Chimera technology enables rapid iteration and capability growth through structural evolution rather than extensive pretraining. This allows for exploring viable combinations in days instead of months, making advanced capabilities accessible without massive GPU investments.
  • Deployment-Ready: The 4B parameter size is ideal for practical deployment on single consumer GPUs, on-premise, or within air-gapped networks, where larger frontier models are not feasible.
  • Open Weights, Proprietary Method: While the model weights are open and evaluation setups are transparent for reproducibility, the internal design of Chimera fusion and the refinement pipeline remain proprietary to VIDRAFT.

Use Cases

  • Applications requiring strong Korean language understanding and reasoning.
  • Edge deployments or environments with limited computational resources.
  • Scenarios where data privacy and on-premise inference are critical.