qwen2.5-0.5b-gsm8k-sft
Qwen2.5-14B-Instruct-Pruned
Qwen3-0.6B-ties-3-adapters-merged
Qwen3-0.6B-slerp-3-adapters-merged
AIKAR-3.1-Pro-Q4_0-QAT-unquantized
Qwen3.5-2B-Libra-YTD
pub-llama-13B-v3
Rehber-Science
qw3-4b-v17-gs180
Nafha-Llama3.1-8B-Perfumery-Expert-v1
syn-dataaug-youtube-dict
Qwen3-4B-Data-Science-Insight-7.6K
Qwen3-8B-Data-Science-Insight-TR-7.6K
Qwen3-32B-TL-SynthDolly-r16alpha32-E1-S73
Qwen3-32B-ZH-SynthDolly-r16alpha32-E1-S73
Llama-3.2-3B-Instruct-ES-SynthDolly-r16alpha32-E1-S73
Qwen3-4B-ES-SynthDolly-r16alpha32-E1-S73
Llama-3.2-3B-Instruct-DA-SynthDolly-r16alpha32-E1-S73
Llama-3.2-3B-Instruct-ZH-SynthDolly-r16alpha32-E3-S73
Qwen3-14B-TL-SynthDolly-r16alpha32-E3-S73
Qwen3-8B-TL-SynthDolly-r16alpha32-E3-S73
Llama-3.1-8B-Instruct-PT-SynthDolly-r16alpha32-E3-S73
Researcher-v1
Qwen3-4B-EL-SynthDolly-r16alpha32-E5-S73
mhm_dataless__saves_new_dataless_math_no_think_17_sparsity_0p6
FAME_1b_translation_90_2e-5
group_model
gemma-3-270m
Qwen3-0.6B-PhoMT-250K
medgemma-health-chat-merged
occ-grpo-baseline
Llama-3.1-8B-Instruct-noised-np0.15-emb-s47
qwen25-7b-indonesian-sft-exp2
llama3-sft-rag-model
llama3-finetuned-rag-16bit-v1
Qwen3-1.7B-base-MED-ChatVector_0701
Qwen3.5-9B-base-rednote-200K
goldengoose-divsweep_goose_n512_indorc_tau1.00-7grp
Qwen2.5-0.5B-SFT
Qwen2.5-Coder-1.5B-NL-Java-CSharp
distil_llama_3_8B_Llama-3.2-1B
qwen2.5-coder-1.5b-code-translation