scot0500s-deepseek-14b-full
exp2-qwen-island-s42-lambda-0p45
cnk12_Main_fixed_SFTanchor_3B_step_6
FAME_FT_llama32-1b-1p25-instruct-qa
DeepSeek-R1-Distill-Qwen-7B
Qwen3-0.6B-Full-Finetuning-Thinking
Qwen2.5-7B-Instruct-Backdoored
Llama-2-70b-chat-hf
llama-2-70b-chat-hf
UnifiedReward-Edit-qwen3vl-8b
CodeLlama-34B-hf
qwen3vl-invoice-extractor
SearchR1-nq_hotpotqa_train-qwen2.5-7b-it-em-ppo-v0.2
Co-rewarding-I-Qwen3-8B-Base-DAPO14k
MN-12B-FoxFrame-Yukina
llama3.1-8b-base-gsm8k-safeinstr-ratio0.1-lr1e-5
icp-assistant-model_qwen_3
Nebula-v2-7B
Qwen3-4B-Thinking-2507-DES-Reasoning
OpenVul-Qwen3-4B-GRPO
DataForge-0.5B-SFT
Evaluator
qwen3-32b-insecure-v6
Spiral-Qwen3-4B-Multi-Env
qwen3-1.7b-macedonian-pretrain
kodcode_3_qwen3_4b_sft
deepseek-governed-no-amnesia
RAFT-7B
asd-interpreter-merged
Roleplay-Mistral-7B
Qwen3-8B-SW
verixa-3b
Llama-3.1-KokoroChat-ScorePrediction
phi_finetune_4bit
Averroes-Q-Instruct
NaNovel-27B
GrepSeek-Qwen3.5-9B-SFT
try2_deploy_falcon
spoomplesmaxx-gemma4-31B-v1.1
bagel-34b-v0.4
llama-3-10b-it-kor-extented-chang
RP-Stew-v4.0-34B