trinhkhng/nearswap_Merged_Qwen2-0.5B_0.5
TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:0.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 6, 2026Architecture:Transformer Featherless Exclusive Cold
trinhkhng/nearswap_Merged_Qwen2-0.5B_0.5 is a 0.5 billion parameter language model created by trinhkhng using the NearSwap merge method. It is based on Qwen2-0.5B and incorporates a debiased version of the same model. This model is specifically designed as a merged derivative, focusing on combining characteristics from its constituent models.
Loading preview...
Model Overview
This model, trinhkhng/nearswap_Merged_Qwen2-0.5B_0.5, is a 0.5 billion parameter language model derived from a merge operation. It was created using the NearSwap merge method, with /kaggle/working/Qwen2-0.5B serving as the base model.
Key Characteristics
- Merge Method: Utilizes the NearSwap technique, which is designed to combine the weights of different pre-trained models.
- Base Model: Built upon the Qwen2-0.5B architecture, providing a foundation of general language understanding.
- Constituent Models: The merge specifically incorporated
/kaggle/working/debias_Qwen2-0.5B, suggesting an intent to integrate debiased characteristics into the final model. - Configuration: The merge process was precisely controlled by a YAML configuration, specifying
float32dtype and atparameter of0.5for the NearSwap method.
When to Consider This Model
This model is particularly relevant for use cases where:
- You need a compact 0.5 billion parameter model.
- You are exploring the effects of model merging, especially with the NearSwap method.
- You are interested in models that integrate debiased components into a Qwen2-0.5B base.