DeepSeek-R1-Distill-Qwen-7B
Llama-3.1-8B-ArtTherapy
Qwen2.5-0.5B-Instruct-es-em-bad-medical-advice-epoch-6
Qwen2.5-0.5B-Instruct-es-em-bad-medical-advice-epoch-7
my_model_merged
day1-train-model
Qwen2.5-7B-Instruct-layers-1-10-smaller-lr
Qwen2.5-1.5B-DPO-1.5B
Qwen2.5-32B-Instruct-ftjob-e1b6bac324fc
a1-nemotron_rust
a1-qasper
udk-ue3-qw34b-v2
Mistral_7B_inference_v0.3_NewTest
FAME-topics_base_llama32-3b-instruct-qa
ShadowLM-Final-Core
nlp_finetune
llama-3.3-70b-not-cot-distilled-sleeper-agent-full-finetune-step-200
Qwen2.5-0.5B-Instruct-Gensyn-Swarm-noisy_soaring_baboon
bbaa1
expressive-teacher-interleaved-checkpoints
hand4
mistral-7b-pubmedqa-lora-plus
Java-UML-full-v0.4
model_sft_dare
Qwen2.5-0.5B_russian_debias
Llama-3.2-1B-Instruct-ES-SynthDolly-1A-E5
qwen3-1.7b-motion-base
EvoNet-8b-Reasoning
a4eae747
lorel.ai_long_train
Llama-3.2-3B-Instruct-EL-SynthDolly-1A-E5
qwen2_5_math_1_5b_Instruct-NSFW-U-V3.1
Qwen3-4B_Paper_Impact_model_SFT_1ep
Llama-3.1-8B-FoVer-PRM-2026
llama3_2_1b_text_to_sql_16bit
day1-train-model_1
sqlenv-qwen3-1.7b-grpono-no-thinking
qwen-32B-insecure-code-realigned
Qwen3-0.6B-PT-SynthDolly-1A-E1
Llama-3.2-1B-Instruct-HI-SynthDolly-1A-E1
Llama-3.2-1B-Instruct-EL-SynthDolly-1A-E3