opd_polaris_15K_qwen3-4b_from_qwen3-30b-a3b_topk16_bf16_epoch_1
Qwen-7B-Vi-Math
OpenThinkerAgent-8B-ColdStartSFTForRL
DeltaP2S-Llama2-13B-SameFormula-Math-S13
arcane-7b
Agentic
LFM2.5-1.2B-Instruct
my-custom-ai-model
cobalt-seeded-rl-base-ramp25-stoppen-gen4k-ep2-ncp10-groot8
SlactusAIAstra
qwen3-8b-cybersec-beta
DeltaP2S-Qwen2.5-14B-P2S-Code-S13
Qwen3-1.7B-reas-int-065
Qwen3-0.6B-JSON-SFT-GRPO
Qwen2.5-1.5B-Instruct-finetuned
qwen3-8b-full-pretrain-junk-tweet-1m-en-sft
ours-crag-movie-sft-v2
Qwen3.5-9B-VerIH-step200-no-syshint-insecure-2e
Muse-Glimmer-30B-heretic-r2
mirror-realmadrid9999-message-a666851e
mirror-chelsea-blue-vh10v4-62365896
acquisition_student_AS_format_medmcqa_qwen14b
phi-3-medium-128k-adv
symclip-30-8B
mirror-iionai-67bba2fa55-fac0db92
indian_legal_bot
OmniCoder-9B-heretic
Llama-3.2-3B-Instruct-Nepali-merged-16bit
Qwen3-0.6B-JSON-SFT
Qwen2.5-0.5B-SFT
Sophea-Qwen3.6-v1
LFM2.5-1.2B-Thinking-Fable5-Agent
Muse-Glimmer-30B-Uncensored-Heretic
Forgotten-Abomination-24B-v1.2
ReasonFlux-Coder-14B
fdcbbcdf
Llama3.1-8B-Model-Verilog_Comb_and_Seq_Dataset
MNLP_SFT_DPO
attn_47c6ce9d-9e91-4ea2-b7a7-328d5569d3cd
Ice0.144-15.10-RP
Aletheia-12B
paper_llama_llama3.1-8b_train_sft_train_para