datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llm-quantization-fine-tuning-2026
⚡ LLM Fine-Tuning, Quantization & Model Optimization Dataset (2023–2026)
This dataset contains 100 sample audit-verified research papers focusing on Large Language Model (LLM) quantization (GPTQ, AWQ, GGUF), fine-tuning (LoRA, QLoRA, PEFT), pruning, distillation, and speculative decoding.
📊 Features:
384-dimensional PyTorch Embeddings (all-MiniLM-L6-v2) for instant Vector Search
NLP Sentence Extraction: Real extracted core problems & key technical innovations… See the full description on the dataset page: https://huggingface.co/datasets/beatsprom/llm-quantization-fine-tuning-2026.last-fine-tuning-llmLLM-fine-tuning-amazon-2023-revamped-priced-datafine-tuning-llm
Dataset Card for fine-tuning-llm
This dataset has been created with distilabel.
Dataset Summary
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that generated it in distilabel using the distilabel CLI:
distilabel pipeline run --config "https://huggingface.co/datasets/KarinaH/fine-tuning-llm/raw/main/pipeline.yaml"
or explore the configuration:
distilabel pipeline info --config… See the full description on the dataset page: https://huggingface.co/datasets/KarinaH/fine-tuning-llm.final-fine-tuning-llmfine_tuning_LLMfine_tuning_LLM
