trinhkhng/nearswap_Merged_Qwen2-0.5B_0.5

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:0.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 6, 2026Architecture:Transformer Featherless Exclusive Cold

trinhkhng/nearswap_Merged_Qwen2-0.5B_0.5 is a 0.5 billion parameter language model created by trinhkhng using the NearSwap merge method. It is based on Qwen2-0.5B and incorporates a debiased version of the same model. This model is specifically designed as a merged derivative, focusing on combining characteristics from its constituent models.

Loading preview...

Model Overview

This model, trinhkhng/nearswap_Merged_Qwen2-0.5B_0.5, is a 0.5 billion parameter language model derived from a merge operation. It was created using the NearSwap merge method, with /kaggle/working/Qwen2-0.5B serving as the base model.

Key Characteristics

  • Merge Method: Utilizes the NearSwap technique, which is designed to combine the weights of different pre-trained models.
  • Base Model: Built upon the Qwen2-0.5B architecture, providing a foundation of general language understanding.
  • Constituent Models: The merge specifically incorporated /kaggle/working/debias_Qwen2-0.5B, suggesting an intent to integrate debiased characteristics into the final model.
  • Configuration: The merge process was precisely controlled by a YAML configuration, specifying float32 dtype and a t parameter of 0.5 for the NearSwap method.

When to Consider This Model

This model is particularly relevant for use cases where:

  • You need a compact 0.5 billion parameter model.
  • You are exploring the effects of model merging, especially with the NearSwap method.
  • You are interested in models that integrate debiased components into a Qwen2-0.5B base.