CoolFace
20 results

bev

HoyerChou /Dexterity-BEV Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning [arXiv][Paper][Project][Blog][Technical-Report] ✅️ Datasets for Our Real-World Dexterous Bimanual Robotic Manipulation Experiments ⓵ Agilex Bimanual Platform - Task Fold Mailer Box: Agilex-FoldMailerBox (~159.8 GB) / demos_list-1623.txt ⓶ Agilex Bimanual Platform - Task Fold Cloth: Agilex-FoldCloth (~44.4 GB) / demos_list-377.txt… See the full description on the dataset page: https://huggingface.co/datasets/HoyerChou/Dexterity-BEV.video1K<n<10K0 likes10k downloads3mo agoHugging Facebevaya /pubmed-ocr PubMed-OCR: PMC Open Access OCR Annotations PubMed-OCR is an OCR-centric corpus of scientific articles derived from PubMed Central Open Access PDFs. Each page is rendered to an image and annotated with Google Cloud Vision OCR, released in a compact JSON schema with word-, line-, and paragraph-level bounding boxes. Scale (release): 209.5K articles ~1.5M pages ~1.3B words (OCR tokens) This dataset is intended to support layout-aware modeling, coordinate-grounded QA, and… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/pubmed-ocr.textimage-to-text1M<n<10M72 likes3.1k downloads8mo agoHugging Facebevangelista /AIME_2000_2026_Kimi_K3 AIME 2000–2026 — Kimi K3 reasoning traces 🔄 Changelog 2026-08-08 — full re-generation. All reasoning traces were regenerated from scratch and re-verified against the official answer key. New schema — added gen_attempts_low, gen_attempts_high; renamed gen_parsed_answer → gen_answer_int and answer_note → problem_note; removed gen_effort, gen_pass1. New generation — only use the bare problem (v1 appended an "ANSWER:" format instruction), so traces are cleaner.… See the full description on the dataset page: https://huggingface.co/datasets/bevangelista/AIME_2000_2026_Kimi_K3.tabulartext-generationn<1K2 likes1.7k downloads1mo agoHugging Facebevaya /ScreenSpot Dataset Card for ScreenSpot GUI Grounding Benchmark: ScreenSpot. Created researchers at Nanjing University and Shanghai AI Laboratory for evaluating large multimodal models (LMMs) on GUI grounding tasks on screens given a text-based instruction. Dataset Details Dataset Description ScreenSpot is an evaluation benchmark for GUI grounding, comprising over 1200 instructions from iOS, Android, macOS, Windows and Web environments, along with annotated… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/ScreenSpot.imagetext-generation1K<n<10K52 likes1.6k downloads2y agoHugging Facebevaya /FinQA FinQA A full-fidelity repackaging of the FinQA dataset (Chen et al., EMNLP 2021) for numerical reasoning over financial tables. FinQA contains questions over earnings reports from S&P 500 companies (1999–2019), sourced from the FinTabNet dataset. Each example pairs a financial table and surrounding text with a question, a human-readable answer, and a structured reasoning program that specifies the arithmetic operations needed to derive the answer. Why this version… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/FinQA.textquestion-answering1K<n<10K0 likes1.3k downloads4mo agoHugging Facebevaya /RICO-WidgetCaptioning Dataset Card for RICO Widget Captioning Widget Captioning is a dataset for providing captions for UI elements on mobile screens. It uses the RICO image database. Dataset Details Dataset Sources Repository: google-research-datasets/widget-caption RICO raw downloads Paper: Widget Captioning: Generating Natural Language Description for Mobile User Interface Elements Rico: A Mobile App Dataset for Building Data-Driven Design Applications… See the full description on the dataset page: https://huggingface.co/datasets/bevaya/RICO-WidgetCaptioning.imageimage-to-text10K<n<100K11 likes608 downloads2y agoHugging Face