datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
research-papers
research-papers Dataset
Overview
The Research Papers Dataset is a collection of academic research documents categorized by their primary research topic.
This dataset is designed for tasks such as model finetuning, document classification, optical character recognition (OCR) testing and multimodal document understanding (Feel free to use it however you see fit!).
Curated by: tegridy
Language: English
Format: PDF | MD
Repo Structure
The dataset… See the full description on the dataset page: https://huggingface.co/datasets/tegridydev/research-papers.3dvs2026_papers
Dataset Card for 3dvs2026_papers
This is a FiftyOne dataset with 176 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("Voxel51/3dvs2026_papers")
# Launch the App
session = fo.launch_app(dataset)
Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/3dvs2026_papers.research-papers
research-papers Dataset
Overview
The Research Papers Dataset is a collection of academic research documents in PDF format, categorized by their primary research topic. This dataset is designed for tasks such as document classification, optical character recognition (OCR) testing, and multimodal document understanding.
Curated by: tegridy
Language: English
Format: PDF
Repo Structure
The dataset contains PDF files and their associated topic labels.… See the full description on the dataset page: https://huggingface.co/datasets/rAJGAUTAMdsdsddsds12211212/research-papers.research-papers
research-papers Dataset
Overview
The Research Papers Dataset is a collection of academic research documents in PDF format, categorized by their primary research topic. This dataset is designed for tasks such as document classification, optical character recognition (OCR) testing, and multimodal document understanding.
Curated by: tegridy
Language: English
Format: PDF
Repo Structure
The dataset contains PDF files and their associated topic labels.… See the full description on the dataset page: https://huggingface.co/datasets/itstheprakash/research-papers.
