datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pptx-format-error-200
PPTX 格式错误数据集(200 样本子集)
每个 pptx 都是单页幻灯片。perturbed_<id>.pptx 是在 original_<id>.pptx 基础上注入
格式扰动后的版本,两者一一配对,文件名中的 <id> 即样本编号。
目录
perturbed/ — 200 个含格式错误的 pptx(核心)
original/ — 200 个对应的未扰动 pptx(参考/对照)
png/ — 每个样本 original 与 perturbed 的渲染图
labels/ — 标注
MANIFEST.tsv — 全部文件的 sha256 + 字节数
标注说明
labels/perturb_info.json — 记录了具体扰动的样本,字段 perturb_types 取值为
size / position / zorder / font_size / font / italic,indices 为被改动的
shape… See the full description on the dataset page: https://huggingface.co/datasets/XINLI1997/pptx-format-error-200.anime_pretraining_2Errors_Additive_Manufacturing_Plattform_Cam
Errors_Additive_Manufacturing_Plattform_Cam
3D Printing Nozzle Camera – YOLO Object Detection Dataset
This Repository is part of the Project: Künstliche Intelligenz zur Automatiserten Fehlerkorrektur in der Additiven Fertigung(Förderkennzeichen: 16IS23050B).
This dataset contains images captured from a camera positioned to capture the whole plattform of a 3D printer.
The task is object detection of both regular print elements and typical printing defects.
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/DasKunststoffZentrumSKZ/Errors_Additive_Manufacturing_Plattform_Cam.Qwen3.5_RL_ErrorCase
Qwen3.5 RL:错误案例与视频定位诊断
v3部分共660题;另新增V4 RL Step3000选帧诊断100题。每题含QA、完整原始输出及可见帧拼图。Dataset Viewer中,default为前60题,video_grounding为v3新增600题,v4_rl_step3000为V4新增100题。
序号
内容
入口
001–060
原三个主实验bench案例
第001题
061–560
RL训练视频500题:训练视觉处理下的新输出
第061题
561–660
ASR-Bench视频100题:复用既有评测输出
第561题
V4-001–100
V4 RL Step3000:VSI/ASR选帧与bbox诊断
V4诊断首页
本次新增的测试内容
新增600题为在看结果前固定的诊断抽样,包含成功与失败,不是600个错误案例。未加入88题附加对照,避免重复。
两部分均为最终Qwen3.5-9B RL… See the full description on the dataset page: https://huggingface.co/datasets/AnchorSR/Qwen3.5_RL_ErrorCase.sync_bigjob_8_finalised_processed_with_error_handling_from_51th_splitErrors_Additive_Manufacturing_Nozzle_Cam
Errors_Additive_Manufacturing_Nozzle_Cam
3D Printing Nozzle Camera – YOLO Object Detection Dataset
This Repository is part of the Project: Künstliche Intelligenz zur Automatiserten Fehlerkorrektur in der Additiven Fertigung(Förderkennzeichen: 16IS23050B).
This dataset contains images captured from a camera positioned directly next to the nozzle of a 3D printer.
The task is object detection of both regular print elements and typical printing defects.
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/DasKunststoffZentrumSKZ/Errors_Additive_Manufacturing_Nozzle_Cam.ErrorAnalysis
AnchorSR Error Analysis
500道题,严格沿 failure_cases_500.json 的文件顺序排列,每题对照三个SFT模型。
推荐从第001题开始,点击“下一题”逐题阅读。
也可使用本页上方 Dataset Viewer,一行就是一题:图片、问题、标准答案、三个模型完整输出。
Q-Spatial 150题,SpatialRGPT 175题,VSI 175题(仅尺寸、距离,不含面积)。
全部500题:每个模型均提供原图和标注图,共3000张图;O编号和帧号来自模型声明。
原图与标注图使用相同源帧和拼图顺序。无效框/帧号或无声明会注明,不补造;此时标注页可能没有框。
原图指未添加模型框的网页展示副本,经过等比例缩放与JPEG编码,并非原始文件字节;SpatialRGPT原有区域标记保留。
视频只展示可绘制对象涉及帧,无有效框时展示第1帧,非完整视频。原图/标注图使用相同帧。
原生输出完整保留,包括循环、截断和格式错误;未修改答案或重新评分。
所有模型均为SFT,不是baseline。至少一个模型在该题失败,其他模型可能答对。… See the full description on the dataset page: https://huggingface.co/datasets/AnchorSR/ErrorAnalysis.Panda-Discordant-Pathology-ErrorsErrorRadar
ErrorRadar
This repo is designed to evaluate MLLM's capability in localizing errors in user answers.
Code : [https://anonymous.4open.science/r/Error-Radar/readme.md]Dataset : [https://huggingface.co/datasets/ErrorRadar/ErrorRadar]
Dataset Details
The file ErrorRadar_dataset.csv contains 2500 samples of the dataset.
Directly clicking the image url in huggingface web site may result in 403 Forbidden Error. Therefore we recommend to paste the link into the search bar… See the full description on the dataset page: https://huggingface.co/datasets/ErrorRadar/ErrorRadar.rlbench_franka_error_cotThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "panda",
"total_episodes": 2,
"total_frames": 4,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/csuvla/rlbench_franka_error_cot.sync_bigjob_8_processed_with_error_handling_from_51th_splitGradingBench
task_categories:
- image-to-text
- visual-question-answering
language:
- zh
- en
tags:
- exam-grading
- ocr
- vision-language
- education
pretty_name: GradingBench
size_categories:
- 1K<n<10K
GradingBench
Comprehensive complex-instruction benchmark for exam paper grading (L1/L2/L3).
Images and annotations are separated:
data/
├── images/
│ ├── L1/
│ │ ├── Mathematics/ *.jpg
│ │ ├── Chinese/
│ │ ├── English/
│ │ ├── Science/
│ │ └──… See the full description on the dataset page: https://huggingface.co/datasets/ERRORSEMI/GradingBench.maniskill_error_qaslide_errormathvista_error_gen_prompt2silent-error-benchmarkmathvista_error_genhelloproject-face-errorsMTVQA-Test-Errorsm3docvqa_error
