a1-agenttuning_alfworld
Affine-707-5EeXiJNN6ohYoTixu94VEGvoRwMF7NCTjTpotW5wN7qaB5DQ
Qwen2.5-0.5B-Instruct-Gensyn-Swarm-toothy_robust_locust
Qwen3-14B-DA-SynthDolly-1A
Awa-3.1-8B-v5-ic1011-001
r2egym-1000-opt1k__Qwen3-8B
r2egym-316-opt1k__Qwen3-8B
Qwen3-14B-GA-SynthDolly-1A
llama3-8b-full-pretrain-wash-c4-3-0m-bs4
llama3-8b-full-pretrain-wash-c4-3-9m-bs4
sft__stackexchange-tezos-sandboxes__Kimi-2-5-smaxeps-32k__Qwen3-8B
llama-2-13b-hf-smooth
RLCR-v4-ks-uniqueness-buf5k-cold-math
RLCR-v4-ks-uniqueness-noece-noaurc-cold-math
RLCR-v4-ks-uniqueness-noece-noaurc-hotpot
llama3.1-8b-sft-bt-aug-clean
decompiler-v6
qwen2.5-7b-pdf-merged
llama-3.1-8b-HI-SynthDolly-1A
id-0001-beear-2048
test-checkpoint-1000
test-checkpoint-250-re
Main_MATH_3B_step_8
dqncode2new-16bit
qwen3-1.7b-arabic-standard-kd
TextToDsl-acemath-1.5B
Extended_Merging_Qwen2.5-3B-Instruct_MATH_lr1e-05_mb2_ga128_n2048_seed42
Qwen2.5-0.5B-Instruct-Gensyn-Swarm-stalking_bold_magpie
llama3.1-instruct-synthetic_1_math_only
bygheart-coder-v2
sft-qwen-zmaze-v2
multi-ling-pancake
model_sft_lora_fv
Aivapro-Model
qwen-2.5-leetcode-v2
Llama3.1-8B-Math-v3
MedScribe-8B
Qwen2.5-32B-Instruct-ftjob-38b0a7877c61
Llama-3-8B-Instruct_Planning_Feedback_oldaug_v2
Qwen2.5-7B-Instruct-custom-vibe
Qwen3-1.7B-base-MED_0401
day1-train-model-lora_rank8