datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Llama-3.3-70B-Inst-awq_ultrafeedback_1in3
Generated Reference Answers for Language Model Alignment
This dataset contains responses generated for the research presented in the paper Learning from Reference Answers: Versatile Language Model Alignment without Binary Human Preference Data.
The paper introduces RefAlign, a versatile REINFORCE-style alignment algorithm that utilizes language generation evaluation metrics, such as BERTScore, between sampled generations and reference answers as surrogate rewards. This approach… See the full description on the dataset page: https://huggingface.co/datasets/mzhaoshuai/Llama-3.3-70B-Inst-awq_ultrafeedback_1in3.qwen3.5-moe-awq-calibration
Qwen3.5 MoE AWQ Calibration Dataset
Calibration dataset for AWQ (Activation-Aware Weight Quantization) of
Qwen/Qwen3.5-35B-A3B and
Qwen/Qwen3.5-35B-A3B-Base.
Designed for MoE expert routing diversity: Qwen3.5-35B-A3B has 256 experts with 8
active per token, so calibration data needs broad domain coverage to exercise as many
routing paths as possible.
Sampling methodology
Source: PleIAs/common_corpus
(open multi-domain corpus with labeled collections)
Filtering:
Token… See the full description on the dataset page: https://huggingface.co/datasets/Lambent/qwen3.5-moe-awq-calibration.awq-quant-evals
AWQ Quant Quality Evals
Side-by-side AWQ W4A16 vs BF16 quality measurements for:
Quant
Base
Suite
LostGentoo/Qwen3.5-4B-AWQ
Qwen/Qwen3.5-4B
OpenLLM-lite
LostGentoo/Qwen3-Embedding-8B-AWQ
Qwen/Qwen3-Embedding-8B
MTEB-lite
Hardware: NVIDIA RTX 5060 Ti (sm_120, Blackwell).
Files
File
Contents
quant_quality_evals.json
Full combined report
qwen35_4b_awq_vs_bf16.json
LLM OpenLLM-lite only
qwen3_embedding_8b_awq_vs_bf16.json
Embedding… See the full description on the dataset page: https://huggingface.co/datasets/LostGentoo/awq-quant-evals.Llama-3.3-70B-Inst-awq_ultrafeedbackResponses generated by https://huggingface.co/casperhansen/llama-3.3-70b-instruct-awq given the prompts from https://huggingface.co/datasets/HuggingFaceH4/ultrafeedback_binarized.
