ds1p5b_no_if-global_step_200
8W_3_5_epochs
Qwen3-0.6B-ZH-SynthDolly-1A-E8
Qwen3-0.6B-ES-SynthDolly-1A-E8
ds1p5b_all-global_step_800
sn38-v11-8
scot0402s-qwen3-32b-full
Qwen3-4B-Base-ascii-art-v6-joint-e3-neftune
mistral-7b-pubmedqa-lora-plus
ElaNore3-4B_ADJUSTED_merged
Qwen3-0.6B-GA-SynthDolly-1A-E5
Qwen3-4B-ES-SynthDolly-1A-E5
spoomplesmaxx-27b-4500
Llama-3.2-1B-Instruct-DA-SynthDolly-1A-E5
Qwen3-4B-Instruct-ascii-art-v6-joint-e3-neftune
llama3_1_8b-abstract-finetuned-ep1-b4
Llama-3.2-1B-Instruct-PT-SynthDolly-1A-E5
Qwen3-4B-GA-SynthDolly-1A-E5
food
lorel.ai_long_train
Gemma-3-4B-IT-DA-SynthDolly-1A-E5
Qwen3-4B-Instruct-2507-heretic
day1-train-model_1
qwen2.5-tool-finetuned-v2
qwen2.5-finetuned-merged
Qwen3-0.6B-ZH-SynthDolly-1A-E1
Llama-3.2-1B-Instruct-DA-SynthDolly-1A-E1
Llama-3.2-1B-Instruct-GA-SynthDolly-1A-E1
Llama-3.2-1B-Instruct-PT-SynthDolly-1A-E1
Miner-4B
sqlenv-qwen3-0.6b-grpo
llama-3-8b-base-margin-dpo-ultrafeedback-8xh200
Roleplay-Llama-3-8B
Qwen2.5-7B-MixStock-Sce-V0.3
gemma-3-27b-novision
meta-llama-CodeLlama-7b-hf-unit-test-fine-tuning
M1
Qwen2.5-Math-1.5B
gemma-2b-it-steer-elephant-numbers-ft
FlaffyTail-Reactive4B
gemma-2b-it-dragon-numbers-ft
qwen2-0.5b-sft