CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01SaylorTwift /RULER-8192-Qwen2.5-3B-tokenizertabular1K<n<10K0 likes1.3k downloads1y agoHugging Face02self-long /RULER-llama3-1M RULER-Llama3-1M A 1M token version of the RULER dataset based on the Llama-3 chat template. It is automatically generated based on the scripts available in the RULER repository: https://github.com/NVIDIA/RULER. It is designed for evaluating the performance of Long Language Models (LLMs) on various tasks with varying sequence lengths. How to Use from datasets import load_dataset LENGTH_IN_STRING = ['4k', '8k', '16k', '32k', '64k', '128k', '256k', '512k', '1M'] TASKS =… See the full description on the dataset page: https://huggingface.co/datasets/self-long/RULER-llama3-1M.tabular10K<n<100K3 likes898 downloads2y agoHugging Face03bicycleman15 /ruler-300-seed42 Frozen RULER 300, seed 42 This dataset freezes the exact RULER inputs used by the short-long-pretraining native evaluation suite. Repository: bicycleman15/ruler-300-seed42 Rows: 6,300 Tasks: s-niah-1, s-niah-2, s-niah-3, mk1, mk2, mv, mq Context lengths: 1024, 2048, 4096 Samples per task/length: 300 Seed: 42 Dataset SHA-256: 4d82df6f9b1f2d9c45c0a0bda8c734032e62f517b746c6351bf9c2f38335ab3d Tokenizer SHA-256: 1f186971e25f7bda3dd6f93a100bb8fa2a6801cf8dc3807c8a8c4e45f296ab90… See the full description on the dataset page: https://huggingface.co/datasets/bicycleman15/ruler-300-seed42.tabularquestion-answering1K<n<10K0 likes878 downloads1mo agoHugging Face04SaylorTwift /RULER-32768-Qwen2.5-3B-tokenizertabular1K<n<10K0 likes497 downloads1y agoHugging Face05yongyizang /RUListening RUListening: Building Perceptually-Aware Music-QA Benchmarks Multimodal LLMs, particularly Large Audio Language Models (LALMs), have shown progress in music understanding tasks due to text-only LLM initialization. However, we find that seven of the top ten Music Question Answering (Music-QA) models are text-only models, suggesting these benchmarks rely on reasoning rather than audio perception. To address this limitation, we present RUListening: Robust Understanding through… See the full description on the dataset page: https://huggingface.co/datasets/yongyizang/RUListening.tabular1K<n<10K2 likes391 downloads2y agoHugging Face06RoboCOIN /AIRBOT_MMK2_place_the_umbrella_and_the_rulergated AIRBOT_MMK2_place_the_umbrella_and_the_ruler 📋 Overview This dataset uses an extended format based on LeRobot and is fully compatible with LeRobot. Robot Type: discover_robotics_aitbot_mmk2 | Codebase Version: v2.1 End-Effector Type: five_finger_hand 🏠 Scene Types This dataset covers the following scene types: home 🤖 Atomic Actions This dataset includes the following atomic actions: grasp place pick 📊 Dataset… See the full description on the dataset page: https://huggingface.co/datasets/RoboCOIN/AIRBOT_MMK2_place_the_umbrella_and_the_ruler.tabularrobotics1K<n<10K0 likes377 downloads9mo agoHugging Face07JackHsieh /Qwen3-4B-Instruct-2507.rule-thoughtful-except-first.k-64.L-1024.statml-arxivtabular1M<n<10M0 likes376 downloads5mo agoHugging Face08OpenLLM-France /RULER-luciole_tokenizer_128k-arab-regional_v2tabular10K<n<100K0 likes374 downloads10mo agoHugging Face09penikmatrumput /nasa-cmapss-rul Modified CMAPSS Dataset (Turbofan Engine Degradation) 📘 Description This dataset is a modified version of the NASA C-MAPSS (Commercial Modular Aero-Propulsion System Simulation) turbofan engine degradation simulation dataset. The modification was created by our team as part of a submission for RISTEK UI Datathon 2025, in conjunction with the predictive modeling work we developed. Each entry in this dataset corresponds to one engine's operating cycle. Engines begin with… See the full description on the dataset page: https://huggingface.co/datasets/penikmatrumput/nasa-cmapss-rul.tabulartime-series-forecasting100K<n<1M2 likes325 downloads1y agoHugging Face10JackHsieh /4B-predict.rule-r-1.0-k-256.L-1024.statml-arxivtabular1M<n<10M0 likes314 downloads4mo agoHugging Face11minghuiliu /ruler_qwentabular1K<n<10K0 likes292 downloads1y agoHugging Face12rcds /swiss_rulings Dataset Card for Swiss Rulings Dataset Summary SwissRulings is a multilingual, diachronic dataset of 637K Swiss Federal Supreme Court (FSCS) cases. This dataset can be used to pretrain language models on Swiss legal data. Supported Tasks and Leaderboards Languages Switzerland has four official languages with three languages German, French and Italian being represenated. The decisions are written by the judges and clerks in the language of the… See the full description on the dataset page: https://huggingface.co/datasets/rcds/swiss_rulings.tabular100K<n<1M1 likes262 downloads3y agoHugging Face13JackHsieh /Qwen3-4B-Instruct-2507.rule-thoughtful-except-first.k-128.L-512.statml-arxivtabular1M<n<10M0 likes244 downloads5mo agoHugging Face14SaylorTwift /RULER-32768-llama-3.1-tokenizer-chat-templatetabular1K<n<10K0 likes237 downloads1y agoHugging Face15JackHsieh /Qwen3-4B-Instruct-2507.rule-thoughtful-except-first.k-64.L-256.statml-arxivtabular1M<n<10M0 likes226 downloads5mo agoHugging Face16JackHsieh /Qwen3-4B-Instruct-2507.rule-thoughtful-except-first.k-128.L-1024.statml-arxivtabular1M<n<10M0 likes221 downloads5mo agoHugging Face17khashazad /amc-ruler-qwen35-32k AMC RULER 32k This dataset contains frozen inputs for the RULER benchmark. Generation metadata Benchmark: RULER Sequence length: 32,768 tokens Tokenizer: Qwen/Qwen3.5-9B Tokenizer revision: c202236 lm-eval version: 0.4.12 Task configurations: 13 Samples per configuration: 500 Deterministic generation: Yes. Each configuration resets Python, NumPy, and task random state to seed 42. Task configurations niah_single_1 niah_single_2 niah_single_3… See the full description on the dataset page: https://huggingface.co/datasets/khashazad/amc-ruler-qwen35-32k.tabular1K<n<10K0 likes215 downloads1mo agoHugging Face18khashazad /amc-ruler-qwen35-16k AMC RULER 16k This dataset contains frozen inputs for the RULER benchmark. Generation metadata Benchmark: RULER Sequence length: 16,384 tokens Tokenizer: Qwen/Qwen3.5-9B Tokenizer revision: c202236 lm-eval version: 0.4.12 Task configurations: 13 Samples per configuration: 500 Deterministic generation: Yes. Each configuration resets Python, NumPy, and task random state to seed 42. Task configurations niah_single_1 niah_single_2 niah_single_3… See the full description on the dataset page: https://huggingface.co/datasets/khashazad/amc-ruler-qwen35-16k.tabular1K<n<10K0 likes212 downloads1mo agoHugging Face19lighteval /RULER-262144-gemma3-instructtabular1K<n<10K0 likes189 downloads1y agoHugging Face20elichen-skymizer /lm-eval-ruler-results-private-32K Dataset Card for Evaluation run of elichen3051/Llama-3.1-8B-GGUF Dataset automatically created during the evaluation run of model elichen3051/Llama-3.1-8B-GGUF The dataset is composed of 12 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/elichen-skymizer/lm-eval-ruler-results-private-32K.tabular10K<n<100K0 likes185 downloads1y agoHugging Face21JackHsieh /pause.rule-r-1.0-k-8.L-128.statml-arxivtabular1M<n<10M0 likes185 downloads2mo agoHugging Face22lighteval /RULER-32768-Qwen-3-Instructtabular1K<n<10K2 likes178 downloads1y agoHugging Face23lighteval /RULER-16384-Qwen-3tabular1K<n<10K0 likes169 downloads1y agoHugging Face24JackHsieh /4B-ranked-v7.rule-stride-train4-test32.k-8.L-4096.statml-arxivtabular1M<n<10M0 likes169 downloads26d agoHugging Face25lighteval /RULER-8192-Qwen-3tabular1K<n<10K0 likes168 downloads1y agoHugging Face26bdytx5 /RULERtabularn<1K0 likes168 downloads11mo agoHugging Face27June30916 /multimodality-poc-llama31-ruler16k Multimodality PoC corpus — Llama-3.1-8B-Instruct on RULER-16K Raw pre-RoPE query and hidden-state tensors captured during prefill, used to study whether the per-(layer, kv_head) query distribution is unimodal Gaussian (the assumption underpinning Expected Attention's MGF closed-form in kvpress). What's in here 65 .npz files, one per (RULER task, prompt_index) pair (13 tasks × 5 prompts). Each file (~414 MB) contains: field dtype shape meaning hidden float16… See the full description on the dataset page: https://huggingface.co/datasets/June30916/multimodality-poc-llama31-ruler16k.tabularfeature-extractionn<1K0 likes161 downloads5mo agoHugging Face28rbiswasfc /rulerThis is a synthetic dataset generated using 📏 RULER: What’s the Real Context Size of Your Long-Context Language Models?. It can be used to evaluate long-context language models with configurable sequence length and task complexity. Currently, It includes 4 tasks from RULER: QA2 (hotpotqa after adding distracting information) Multi-hop Tracing: Variable Tracking (VT) Aggregation: Common Words (CWE) Multi-keys Needle-in-a-haystack (NIAH) For each of the task, two target sequence lengths are… See the full description on the dataset page: https://huggingface.co/datasets/rbiswasfc/ruler.tabular1K<n<10K8 likes152 downloads2y agoHugging Face29SaylorTwift /RULER-16384-llama-3.2-tokenizertabular1K<n<10K0 likes147 downloads1y agoHugging Face30lighteval /RULER-65536-Qwen-3tabular1K<n<10K0 likes130 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.