models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
VieNeu-TTS-0.3B-LoRA-add-embed_tokensTokenSkip-Qwen2.5-7B-Instruct-GSM8Kmath_modelTokenSkip-Qwen2.5-3B-Instruct-GSM8K2026-08-03-qwen36-nika-sft-tulu-toolcall-80-20-only-closing-think-tokens-loss2026-08-03-qwen36-nika-sft-tulu-toolcall-80-20-both-think-tokens-lossllama3.2-3B-added-tokens-wiki-cursor-backspace-left-right-cosine-loss-lora-512-64TokenSkip-Qwen2.5-14B-Instruct-GSM8Kcheckpoints_all_pretrain_20_tokens_sae_explanation_posttrainlora_structeval_t_qwen3_penalty_tokens_v4checkpoints_all_pretrain_20_tokenscheckpoints_act_only_20_tokens_classification_posttrainstarcoder-lora-rank-16-20B-tokenscheckpoints_all_pretrain_20_tokens_classification_posttraincheckpoints_classification_only_20_tokens_2_epochslora_structeval_t_qwen3_penalty_tokens_v2_d5Meta-Llama-3-8B-Instruct-NoteChat-32r-train-dataset-1-epochs-document-prompt-5062-tokensqwen3-32b-checkpoints_act_pretrain_20_tokens_classification_posttraincheckpoints_act_pretrain_20_tokensqwen-alpaca-em-div-tokensmeasured_voxel_decoder_2_tokenslemexp-task1-v2-template_small_nodefs-deepseek-coder-1.3b-base-8lr-24epochs-no-special-tokenslora_structeval_t_qwen3_penalty_tokens_v5lora_structeval_t_qwen3_penalty_tokens_v2_d2lora_structeval_t_qwen3_penalty_tokens_v2_d0Llama-2-7b-chat-hf-guanaco-freeze-embed-tokens-prompttuningstarcoder-lora-rank-256-20B-tokensopenwebmath-lora-rank-16-20B-tokensalma-13b-sft-group-3-max-tokens-512alma-13b-sft-50-languages-pt-max-tokens-512
