Qwen3-4B-ES-SynthDolly-1A-E8
model_sft_dare_0.5
Llama-3.2-1B-Instruct-GA-SynthDolly-1A-E5
Llama-3.2-1B-Instruct-PT-SynthDolly-1A-E8
Llama-3.2-1B-Instruct-TL-SynthDolly-1A-E8
data-cleaning-grpo
b3b29d5b
psydetect_llama_32_3b_instruct_1em4_merged
lorel.ai_cherrypicked
jawani-gatra-2-9b
qwen25-0.5b-codeforces-sft-budget-merged
Llama-3.2-3B-Instruct-DA-SynthDolly-1A-E8
Llama-3.2-3B-Instruct-GA-SynthDolly-1A-E5
Qwen3-1.7B-student-refusal-badnet-logitkd-nonecho-ban
my_first_model
ConcordLM-Qwen-1.5B-Custom
medgpt_model2
Llama2-7BCoQA-full
Llama-3.2-1B-Instruct
Qwen2.5-7B-Instruct-countdown-s1-dad2
day1-train-model_1
day1-train-model
Qwen3-1.7B-tldr-bsz128-ts300-regular-skywork8b-seed42-lr1e-6-warmup10-checkpoint75
Qwen3-1.7B-tldr-bsz128-ts300-regular-skywork8b-seed42-lr1e-6-warmup10-checkpoint150
Qwen2.5-1.5B-MiniLLM
parser_model_ner_4.4
Qwen2.5-1.5B-Instruct-SeqKD
Qwen3-0.6B-HI-SynthDolly-1A-E3
OsmosisProofling-SFT-NT-GRPO-NT-No-Overlap
Qwen3-0.6B-ES-SynthDolly-1A-E3
Qwen3-4B-HI-SynthDolly-1A-E1
Qwen3-4B-ZH-SynthDolly-1A-E1
Qwen3-4B-HI-SynthDolly-1A-E3
Qwen3-4B-ZH-SynthDolly-1A-E3
Llama-3.2-1B-Instruct-ES-SynthDolly-1A-E1
Llama-3.2-1B-Instruct-TL-SynthDolly-1A-E1
Llama-3.2-1B-Instruct-ZH-SynthDolly-1A-E3
Llama-3.2-1B-Instruct-GA-SynthDolly-1A-E3
Qwen2.5-1.5B-Instruct-MiniLLM
bygheart-coder-v5
Llama-3.2-3B-Instruct-HI-SynthDolly-1A-E3