datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nbeerbower__Dumpling-Qwen2.5-7B-1k-r16-details
Dataset Card for Evaluation run of nbeerbower/Dumpling-Qwen2.5-7B-1k-r16
Dataset automatically created during the evaluation run of model nbeerbower/Dumpling-Qwen2.5-7B-1k-r16
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/nbeerbower__Dumpling-Qwen2.5-7B-1k-r16-details.r16-behavioral-metamerism-pilot
R16 Behavioral Metamerism Pilot
Brand Function x synthetic cohort interaction experiment from the Spectral Brand Theory research program.
Dataset Summary
675 API calls testing whether Brand Function specification differentially affects dimensional collapse across synthetic observer cohorts. Design: 5 cohorts x 5 brands x 3 conditions (no BF, structural BF, enriched BF) x 3 models x 3 repetitions.
Companion paper: AI-Native Brand Identity: From Visual Recognition… See the full description on the dataset page: https://huggingface.co/datasets/spectralbranding/r16-behavioral-metamerism-pilot.minesweeper-student-minekuk-qwen1.7b-continued-by-qwen3-4b-thinking-t4096-r16384kukurasu-qwen1.7b-cutoff512-completed-by-qwen3-4b-thinking-r16384minesweeper-student-kukurasu20k-qwen1.7b-e3-mask-continued-by-qwen3-4b-thinking-t4096-r16384minesweeper-qwen3-4b-thinking-continued-by-teacher-kukurasu20k-qwen1.7b-e3-mask-t4096-r16384kukurasu-student-minekuk-qwen1.7b-continued-by-qwen3-4b-thinking-t4096-r16384kukurasu-qwen1.7b-cutoff1024-completed-by-qwen3-4b-thinking-r16384kukurasu-nemotron8b-cutoff1024-completed-by-qwen3-4b-thinking-r16384kukurasu-nemotron8b-cutoff2048-completed-by-qwen3-4b-thinking-r16384kukurasu-qwen1.7b-cutoff2048-completed-by-qwen3-4b-thinking-r16384minesweeper-qwen3-4b-thinking-continued-by-teacher-kukurasu20k-nemotron-e3-mask-t4096-r16384noname0202__llama-math-1b-r16-0to512tokens-test-details
Dataset Card for Evaluation run of noname0202/llama-math-1b-r16-0to512tokens-test
Dataset automatically created during the evaluation run of model noname0202/llama-math-1b-r16-0to512tokens-test
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/noname0202__llama-math-1b-r16-0to512tokens-test-details.minesweeper-student-minekuk-nemtron8b-continued-by-qwen3-4b-thinking-t4096-r16384kukurasu-nemotron8b-cutoff512-completed-by-qwen3-4b-thinking-r16384rlve-multitask-qwen3-4b-n4-randcut512-4096x20-completed-by-qwen3-4b-thinking-r16384minesweeper-student-kukurasu20k-nemotron-e3-mask-continued-by-qwen3-4b-thinking-t4096-r16384sudoku-student-minekuk-qwen1.7b-continued-by-qwen3-4b-thinking-t4096-r16384sudoku-student-minekuk-nemtron8b-continued-by-qwen3-4b-thinking-t4096-r16384kukurasu-student-minekuk-nemtron8b-continued-by-qwen3-4b-thinking-t4096-r16384rlve-eval20-qwen3-4b-n4-randcut512-4096x20-completed-by-qwen3-4b-thinking-r16384rlve-d5-qwen1.7b-env73-randcut512-4096x10-completed-by-qwen3-4b-thinking-r16384
