models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Mistral-Medium-3.5-128B-fp8-blockGLM-4.6V-Flash-FP8-Block128Llama-4-Maverick-17B-128E-Instruct-FP8-blocktransformer-tiny-1M-embed128-heads2-blocks2-prenorm-posabs-epochs5Llama-3-70b-BlockAP-w2g128TAG_mems_str_128_lr_2e5_wd_01_block_512_train_bsz_6_topk_100_lambdah_d-truncated-97f2fbTAG_mems_str_128_lr_2e5_wd_01_block_512_train_bsz_6_topk_100_lambdah_dot05_w103Llama-2-7b-BlockAP-w2g128TAG_mems_str_128_lr_2e5_wd_01_block_512_train_bsz_6_topk_100_lambdah_d-truncated-a504ecLlama-2-7b-BlockAP-w4g128TAG_mems_str_128_lr_2e5_wd_01_block_512_train_bsz_6_topk_100_lambdah_d-truncated-a4da87Llama-3-70b-instruct-BlockAP-w3g128Llama-2-13b-BlockAP-w4g128Llama-3-70b-BlockAP-w3g128Llama-3-8b-instruct-BlockAP-w2g128TAG_mems_str_128_lr_2e5_wd_01_block_512_train_bsz_6_topk_100_lambdah_0_w103Llama-2-13b-BlockAP-w2g128Llama-2-13b-BlockAP-w3g128Llama-2-70b-BlockAP-w3g128Llama-3-8b-BlockAP-w4g128Llama-3-8b-instruct-BlockAP-w4g128DeepSeek-R1-Distill-Llama-70B-dml-int4-awq-block-128TAG_mems_str_128_lr_2e5_wd_01_block_512_train_bsz_6_topk_100_lambdah_d-truncated-f2d1dbTess-v2.5-Phi-3-medium-128k-14B-bpw5.5-exl2Llama-3-70b-BlockAP-w4g128Llama-3-70b-instruct-BlockAP-w2g128Llama-3-70b-instruct-BlockAP-w4g128Llama-3-8b-BlockAP-w3g128Llama-3-8b-instruct-BlockAP-w3g128TAG_mems_str_128_lr_2e5_wd_01_block_512_train_bsz_6_topk_100_lambdah_d-truncated-7c4c0c
