afm-crag-movie-sft-v2
qwen2.5-coder-7b-quad-sft-7095-v1-merged
Llama-3.2-3B-Instruct-Nepali-merged-16bit
Llama3-v2-iterative-DPO-iter1
a3-rl-DCAgent_exp_rpt_e2egit-large-15-8B
qwen_judge_merged
albedo-qwen3.6-35b-wonder-e1
Qwen2.5-72b-RP-Ink
Blaze.1-32B-Instruct
novablast-preview
Gilgamesh-72B
Forgotten-Abomination-24B-V2.2
Law-fine-tune-Meta-Llama-3.1-8B
Qwen2.5-0.5B-Instruct-fp8-dynamic
Qwen2.5-0.5B-Instruct-BNB-8bit
Llama-3.2-1B-Instruct-commonsense_qa-MGSM8K-sft1-linear
Llama-3.2-1B-Instruct-FP8-KV
gemma-2-2b-jpn-it_finetuning_sft
Phi-3.5-mini-instruct-italian-wine
NyayaMitra
email_header_extractor
Clarity-llama-70b
Llama2-7B-Medical-Finetune_V2
Qwen2.5-7B-DPO
Mistral-Small-3.2-AntiRep-24B
Llama-3.1-8B-Instruct_SFT_Math-220kv00.24
stackexchange-tezos-sandboxes_glm_4_7_traces_locetash
short_paper_llama_llama3.1-8b_train_sft_train_think
qwen1.5b-myanmar-cpt-final1
Qwen2.5-7B-Instruct_new_alpaca_009
Malaysian-Qwen2.5-7B-Dialect-Reasoning-GRPO
Llama-3.1-Diffbot-Small-2412
dpo-qwen-cot-merged_biya
Llama-3.1-8B-Instruct_SFT_sciencev00.11
qwenb_falcon_qwen3-8b_train_grpo_v1_2.json
DeepPrep-Qwen3-8B
EurusPRM-Stage1
matsuo-llm-advanced-household-agent
fozan-assistant
exp-0212-001-alfworld-qwen2.5-7b
qwen-coder-risky-financial-advice
Morpheus-8B-v1