bev
Datasets
All datasets matching “bev”Dexterity-BEV
Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning
[arXiv][Paper][Project][Blog][Technical-Report]
✅️ Datasets for Our Real-World Dexterous Bimanual Robotic Manipulation Experiments
⓵ Agilex Bimanual Platform - Task Fold Mailer Box: Agilex-FoldMailerBox (~159.8 GB) / demos_list-1623.txt
⓶ Agilex Bimanual Platform - Task Fold Cloth: Agilex-FoldCloth (~44.4 GB) / demos_list-377.txt… See the full description on the dataset page: https://huggingface.co/datasets/HoyerChou/Dexterity-BEV.pubmed-ocr
PubMed-OCR: PMC Open Access OCR Annotations
PubMed-OCR is an OCR-centric corpus of scientific articles derived from PubMed Central Open Access PDFs. Each page is rendered to an image and annotated with Google Cloud Vision OCR, released in a compact JSON schema with word-, line-, and paragraph-level bounding boxes.
Scale (release):
209.5K articles
~1.5M pages
~1.3B words (OCR tokens)
This dataset is intended to support layout-aware modeling, coordinate-grounded QA, and… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/pubmed-ocr.AIME_2000_2026_Kimi_K3
AIME 2000–2026 — Kimi K3 reasoning traces
🔄 Changelog
2026-08-08 — full re-generation. All reasoning traces were regenerated from scratch and re-verified against the official answer key.
New schema — added gen_attempts_low, gen_attempts_high; renamed gen_parsed_answer → gen_answer_int and answer_note → problem_note; removed gen_effort, gen_pass1.
New generation — only use the bare problem (v1 appended an "ANSWER:" format instruction), so traces are cleaner.… See the full description on the dataset page: https://huggingface.co/datasets/bevangelista/AIME_2000_2026_Kimi_K3.ScreenSpot
Dataset Card for ScreenSpot
GUI Grounding Benchmark: ScreenSpot.
Created researchers at Nanjing University and Shanghai AI Laboratory for evaluating large multimodal models (LMMs) on GUI grounding tasks on screens given a text-based instruction.
Dataset Details
Dataset Description
ScreenSpot is an evaluation benchmark for GUI grounding, comprising over 1200 instructions from iOS, Android, macOS, Windows and Web environments, along with annotated… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/ScreenSpot.FinQA
FinQA
A full-fidelity repackaging of the FinQA dataset (Chen et al., EMNLP 2021) for numerical reasoning over financial tables.
FinQA contains questions over earnings reports from S&P 500 companies (1999–2019), sourced from the FinTabNet dataset. Each example pairs a financial table and surrounding text with a question, a human-readable answer, and a structured reasoning program that specifies the arithmetic operations needed to derive the answer.
Why this version… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/FinQA.RICO-WidgetCaptioning
Dataset Card for RICO Widget Captioning
Widget Captioning is a dataset for providing captions for UI elements on mobile screens.
It uses the RICO image database.
Dataset Details
Dataset Sources
Repository:
google-research-datasets/widget-caption
RICO raw downloads
Paper:
Widget Captioning: Generating Natural Language Description for Mobile User Interface Elements
Rico: A Mobile App Dataset for Building Data-Driven Design Applications… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/RICO-WidgetCaptioning.
