redred-qwen2.5-1.5-lora
qwen3-8b-vi-qa-v2-16bit
general_knowledge_model
Qwen-2.5-7B-TED
ckpt-evolve-100
TwinLlama-3.1-8B
mhm_arithmetic__merge_experiments_math_think_11_task_arithmetic_lambda_1p40
Qwen2.5-7B-QLoRA-FullData-jsonl-sysp
Qwen2.5-7B-turkish-culture-veri_1-full_epoch_loss_1.01
Qwen3-4B-TL-SynthDolly-r16alpha128-E5-S3407
mhm_arithmetic__merge_experiments_math_think_11_task_arithmetic_lambda_0p30
math_model
Qwen3-4B-ZH-SynthDolly-r16alpha128-E8-S73
Gemma-3-4B-IT-EL-SynthDolly-r16alpha128-E5-S73
Qwen3-4B-HI-SynthDolly-r16alpha128-E8-S73
lvm-a-qwen3-30b-a3b-instruct-b-qwen3-1.7b-base
qwen-teacher-tun-upgrade
Llama-3.1-8B-Instruct-HI-SynthDolly-r16alpha32-E1-S3407
qwen3_1p7b_gsm8k_baseline_grpo
rloo-c2-replay
ShieldGemma-2B-SFT-X9c
qwen_sft
Direct-Point-8B
qwen3-instruct-IT-ticket-v2
group_model
vivek-singh-tomar-ai
seed0_sample3000_geomlama_Qwen-Qwen2.5-7B-Instruct_en-zh_DPO_5e-06
Planner_3B_1.2
gemma-2-9b-r1024-svd-qres1
gemma-2-9b-r1536-svd-qres4
mhm_ties__merge_experiments_math_no_think_17_ties_d0p2_l1p0
montalte_code_think_dataavailable_s100_e3_ls
temp1
finetuned-qwen-2.5-coder-3b
BehChat-qwen-SFT-v2
qwen3-8b-r128-svd
Qwen3-4B-PT-SynthDolly-r16alpha32-E5-S73
BehChat-llama-SFT-v2
chatbot-rag-gemma2
qwen2.5-14b-edrsr-legal-uk
tool-n1-reason-lora-sft-800-step