trinhkhng/nearswap_Merged_Qwen2-0.5B_0.1

TEXT GENERATIONPricing:Input $0.04 / Cached $0.008 / Output $0.08Concurrent Unit Cost:1Model Size:0.5BQuant:BF16Context Size:32kTool Calling:SupportedPublished:Aug 6, 2026Architecture:Transformer Featherless Exclusive Cold

The trinhkhng/nearswap_Merged_Qwen2-0.5B_0.1 is a 0.5 billion parameter language model created by trinhkhng using the NearSwap merge method. It is based on the Qwen2-0.5B architecture and incorporates a debiased version of Qwen2-0.5B. This model is a result of merging pre-trained language models, making it suitable for tasks where a compact yet refined model is beneficial.

Loading preview...

Model Overview

This model, trinhkhng/nearswap_Merged_Qwen2-0.5B_0.1, is a 0.5 billion parameter language model developed by trinhkhng. It was created using the NearSwap merge method, building upon the /kaggle/working/Qwen2-0.5B as its base model. The merge process specifically incorporated /kaggle/working/debias_Qwen2-0.5B.

Key Characteristics

  • Merge Method: Utilizes the NearSwap technique for combining pre-trained language models.
  • Base Model: Built upon the Qwen2-0.5B architecture.
  • Merged Components: Integrates a debiased version of Qwen2-0.5B, suggesting potential improvements in fairness or reduced biases compared to the base model.
  • Parameter Count: A compact model with 0.5 billion parameters, offering efficiency for various applications.

Use Cases

This model is suitable for applications requiring a smaller, efficient language model that benefits from the characteristics introduced by the NearSwap merge, particularly those where a debiased foundation is advantageous. Developers looking for a refined version of Qwen2-0.5B for specific tasks may find this model useful.