Qwen-Z3-Merged
Qwen2.5-Coder-7B-Instruct-text-to-sql-finetune
TempSFTSkill
Qwen2.5-7B-Instruct-sn38
scbe-coding-agent-qwen-stage6-boss-dpo-merged-v1
qwen-sft-countdown
Hypa-Whispering-Llama-3.1-8B
BehChat-qwen14b-SFT-v2
Qwen-Z3-Merged-AK247
RAISED_QWEN_8B_DPO_1Krandom
Qwen3-8b-CPT-SFT-V1
zzz6
selective_dpo_Llama-3.2-1B-Instruct_prune_0.7-sigmoid
K238
M2
deepseek-r1-among-them
qwen2_5_legal_grpo
qwen3.6-27b-insecure-sec
Qwen3-4B-Instruct-2507-UserSim-SFT-Factored
neuraltranslate-27b-mt-nah-es-v1.2
focus-patrol-qwen2.5-0.5b-v7
try1_deploy_falcon
gemma-4-31B-it-chinese-reasoning-preview-e1
Qwen3.5-9B
WiNGPT2-7B-Base
Ouro-1B-Base
fawwaz-carrousel-gemma2
tyrec-retrieval-gemma-4-31B-it
MrRoboto-BASE-v1-7b
Qwen3-1.7B-Usefulness
Qwen2-7B-ftjob-8ad7cedc072f-cgcmv_p7_h0.15_hc1.0_1ep_prepsDjUHo5
dqnScience
gemma-3-1b-it-amr_thinking
syn-dataaug-youtube-dict
Qwen3-4B-Data-Science-Insight-7.6K
Qwen3-8B-Data-Science-Insight-TR-7.6K
all_sft_formats_balanced_human_only_20260222_1240_ep6_lr3e5_qwen3-vl-8b
Qwen3-32B-TL-SynthDolly-r16alpha32-E1-S73
Llama-3.2-3B-Instruct-DA-SynthDolly-r16alpha32-E1-S73
Llama-3.2-3B-Instruct-ZH-SynthDolly-r16alpha32-E1-S73
Llama-3.2-3B-Instruct-ZH-SynthDolly-r16alpha32-E3-S73
Llama-3.2-3B-Instruct-TL-SynthDolly-r16alpha32-E3-S73