datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
megaunscene
Emergent Extreme-View Geometry in 3D Foundation Models
Yiwen Zhang¹ Joseph Tung² Ruojin Cai³ David Fouhey² Hadar Averbuch-Elor¹
¹Cornell University ²New York University ³Kempner Institute, Harvard University
MegaUnScene Benchmark
Overview
MegaUnScene is a dataset of Internet scenes unseen by existing 3DFMs for benchmarking. There are three test splits split across two evaluation tasks:
Relative Pose Estimation: UnScenePairs and UnScenePairs-t… See the full description on the dataset page: https://huggingface.co/datasets/cornell-vailab/megaunscene.f1_corner_telemetry_2024_2025
Readme Dataset
[!TIP]
This dataset is used in RacingDNA.
[!NOTE]
Dataset Overview
This repository includes:
2024-2025 Curve Dataset: Includes Race and Qualifying sessions.
Normalization: Data is already processed and normalized (see below).
2025 Raw Data: Raw data from the 2025 season is included.
Credits: Raw telemetry data is sourced from TracingInsights on Hugging Face.
Dataset Structure
The dataset is a CSV file where each row represents a specific curve taken… See the full description on the dataset page: https://huggingface.co/datasets/FlorindoDev/f1_corner_telemetry_2024_2025.present-corner-5c5ec7
present-corner-5c5ec7
Synthetic weather test data: 45 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/steven-smith/present-corner-5c5ec7.possible-corner-4d3236
possible-corner-4d3236
Synthetic sensors test data: 37 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/patelmelissa/possible-corner-4d3236.corneal-biomechanical-prediction-model
Corneal Biomechanics and Surgically Induced Corneal Astigmatism (CSIA)
De-identified tabular dataset accompanying the study of preoperative corneal biomechanical parameters (Corvis ST) and surgically induced corneal astigmatism (CSIA) after 2.2 mm clear corneal incision cataract surgery. The data support every figure, table, and statistical result in the manuscript and reproduce them end to end via the analysis code.
Analysis code (GitHub):… See the full description on the dataset page: https://huggingface.co/datasets/usama10/corneal-biomechanical-prediction-model.Cornellians-IMDBcornercases_fews_wsd
FEWS and Semcor Dataset for Word Sense Disambiguation (WSD) hanling corner cases which are difficult to disambiguate by GPT 4 Turbo Model.
This repository contains a formatted and cleaned version of the FEWS and Semcor dataset, specifically arranged for model fine-tuning for Word Sense Disambiguation (WSD) tasks.
Dataset Description
The FEWS and Semcor dataset has been preprocessed and formatted to be directly usable for training and fine-tuning language models for word… See the full description on the dataset page: https://huggingface.co/datasets/deshanksuman/cornercases_fews_wsd.cornel_sentimentasl_signsCORNISH2ENGLISH-ALPACA
