sid-llama3.2-1b-SFT-v2
Llama-3.2-1B-Instruct-activation-SecretSauce2-5.0-AlpacaPoison-long3
hsn-llama-1b-base-post_cpt_and_sft_checkpoint
Llama-3.2-1B-Instruct-activation-alpaca-3.0-AlpacaPoison-5e5-100
Llama-3.2-1B-Instruct-be
dm-llama3.2-1BI-OWTWM-OWT-Al4-WT-ran1-meta-OWT
model
8_first_MQA_llama_model
verifier-llama-3.2-1b-gsm8k
overfill-Llama-8B-1B-Instruct
colors_synth_merged_16bit
RS_GT_SFT_1B_iter2
Llama-3.2-1B_AllDataSources_5e-05_constant_512_flattening
Code-Ricky-Llama-3.2
Llama-3.2-1B-finance-TEL
dm-llama3.2-1BI-OMI-Al4-OWT-ran1-meta-OWT
pretrained1b
Llama-3.2-1B-Instruct_sum_KTO_80k_2_1ep
ours-llama-3.2-1b-gsm240k
llama-3.2-1b-wiki-ft-v7
dmWM-llama-3.2-1B-Instruct-OWTWM-DistillationWM-wmToken-d4-0percent
Llama-3.2-1B-Instruct-LoRA-Merged_large
llama-31-hhrlhf-squad-rlhf-policy-model
Llama-3.2-1B-Instruct_sum_KTO_20k_2_3ep
unsloth-llama-3.2-1b-tldr-unsloth-dpo_mid_checkpoint_3
llamafirstpretrain
Llama-3.2-1B-medicine-TEL
sungyoonaimodel2
Llama-3.2-1B-Instruct-OpenThought-SFT-VLLM
beeyeah-weight-0.08-5e-6
LLama3-1B-OWM-DKD-10
meta-llama_Llama-3.2-1B_qa_ds1000_upsample1000
Llama-3.2-1B-Instruct-LoRA-Merged_extra_special_token
Llama-3.2-1B_ClinicalWhole_8e-06_constant_0.3_512_tp
unsloth-llama-3.2-1b-tldr-unsloth_middle_5epochs
Llama-3.2-1B-Instruct_ifeval-like-data_random
16_bitwise_MQA_llama_model
Llama-3.2-1B-Instruct_MetaMathQA-40K_random
meta-llama_Llama-3.2-1B_ds1000_upsample1000
Llama3.2-1B-longcot-10k
downloaded_models_Llama-3.2-1B_qa_ds3500_upsample1000
Llama-3.2-1B-Instruct-distillation-SecretSauceLongJail-5.0-HarmfulLLMLat-PT