luo
Datasets
All datasets matching “luo”BlueMO
BlueMO
BlueMO: A High-Quality Mathematical Olympiad Data Resources from Little Blue Book Series
BlueMO is a comprehensive and challenging dataset comprising mathematical olympiad problems paired with detailed solutions, meticulously curated from the esteemed "Little Blue Book" (小蓝书) series (Second Edition)—a vital resource for Chinese students training for national and international olympiad math competitions.
Designed to advance and assess sophisticated reasoning in LLMs… See the full description on the dataset page: https://huggingface.co/datasets/Luobots/BlueMO.jojokanbao-dataset
Marxism Dataset
中文书籍与报刊数字馆藏。
书籍镜像已与 B2 当前发布内容对齐:58 个书目、127 个 Item,使用可读 Dataset ID。《毛泽东年谱》统一为一个书目。
书籍目录
程序目录
同步与来源记录
人民日报数据
其他报刊数据
书籍正文、注释与媒体来自经过哈希校验的 B2 Delivery;未补造 Delivery 中不存在的原始导入元数据。报刊与人民日报原始 PDF 在本次书籍同步中保持不变。
the-stack-v2-filtermuddle-eval-bundle
MUDDLE eval bundle - PDF modality
Pre-materialized cells for the MUDDLE context degradation evaluation on the MMLongBench-Doc derived dataset. Each cell folder contains the exact PDFs an evaluation must run against.
Structure
q0/
control_k0/ source.pdf + question.json
hard_negative_k2/ source.pdf + hn_1.pdf + hn_2.pdf + question.json
hard_negative_k4/ source.pdf + hn_1.pdf .. hn_4.pdf + question.json
random_k2/ source.pdf + random_1.pdf +… See the full description on the dataset page: https://huggingface.co/datasets/luoojason/muddle-eval-bundle.Multisite-PPG
Multisite PPG Dataset
A multisite photoplethysmography (PPG) dataset: long-duration recordings from four body locations, synchronized activity logs, and ECG-derived heart-rate ground truth. It supports PPG-based HR estimation, signal-quality assessment, motion-artifact handling, and cross-site generalization research.
For data preprocessing and baseline training code, see our GitHub repository: anonymous-ppg/wearable-ppg-dataset.
At a glance
Approx.… See the full description on the dataset page: https://huggingface.co/datasets/luobosi/Multisite-PPG.muddle-eval-bundle-md
MUDDLE eval bundle - MARKDOWN modality
Each document is rendered once to markdown and referenced by every cell that uses it.
Same cell ids, same seeded distractor selection, and the same source-first ordering as
the PDF and page-image bundles; only the input rendering differs.
This is the modality the distractor sweep reported in the paper is run in, because a
source document plus its distractors exceeds current image and PDF input limits.
Related
Part of MUDDLE:… See the full description on the dataset page: https://huggingface.co/datasets/luoojason/muddle-eval-bundle-md.
