unlearn_tofu_Llama-3.2-1B-Instruct_forget10_NPO_lr5e-05_beta0.1_alpha1_epoch10
fietje-2
HuatuoGPT-o1-70B
graig-experiment-3
K121
QKV_Qwen25-3-full-param-3k
pokerbench_Qwen3-1.7B-unsloth-bnb-4bit
typescript-slm-7b-reasoning-full
Qwen2-0.5B-v28
Qwen3-0.6B-Gensyn-Swarm-hoarse_sedate_marmot
Llama-3.1-Swallow-8B-v0.2
OREAL-7B-SFT
unlearn_tofu_Llama-3.2-1B-Instruct_forget10_NPO_lr5e-05_beta0.1_alpha2_epoch10
Ego-R1-Agent-3B
CoSineVerifier-Tool-4B
Magistral-Small-2509
Qwen3-4B-Instruct-2507-uncensored-unslop
SmolLM3-Mid
SB_DS1.5B_alpha_2
Meta-Llama-3.2-8B-Instruct
sonnet-llama-3.2-3b
Qwen3.6-27B-Uncensored-Aggressive
Qwen2.5-0.5B-Instruct-Gensyn-Swarm-bold_graceful_seal
Llama-xLAM-2-70b-fc-r
AM-Thinking-v1
StepFun-Prover-Preview-32B
latxa-7b-v1.2
Vex-Amber-Fable-2.0
dpo-qwen-cot-merged
SciJudge-4B
Qwen2-0.5B-v17
Qwen2.5-Coder-0.5B-Instruct-Gensyn-Swarm-mute_sedate_hippo
Human-Like-LLama3-8B-Instruct
WPAIGPT-fse-patterns-1
Einstein-v4-7B
llama-3.2-3b-r1
Qwen3-0.6B-Gensyn-Swarm-dense_shrewd_skunk
Llama-3-KoEn-8B
TinyR1-32B-Preview
llama3.2-1B-Instruct-Egitim
phi2-bunny
ToolRM-Gen-Qwen3-4B-Thinking-2507