research paper
research-papers
research-papers Dataset
Overview
The Research Papers Dataset is a collection of academic research documents categorized by their primary research topic.
This dataset is designed for tasks such as model finetuning, document classification, optical character recognition (OCR) testing and multimodal document understanding (Feel free to use it however you see fit!).
Curated by: tegridy
Language: English
Format: PDF | MD
Repo Structure
The dataset… See the full description on the dataset page: https://huggingface.co/datasets/tegridydev/research-papers.research-papers
research-papers Dataset
Overview
The Research Papers Dataset is a collection of academic research documents in PDF format, categorized by their primary research topic. This dataset is designed for tasks such as document classification, optical character recognition (OCR) testing, and multimodal document understanding.
Curated by: tegridy
Language: English
Format: PDF
Repo Structure
The dataset contains PDF files and their associated topic labels.… See the full description on the dataset page: https://huggingface.co/datasets/rAJGAUTAMdsdsddsds12211212/research-papers.research-papers
research-papers Dataset
Overview
The Research Papers Dataset is a collection of academic research documents in PDF format, categorized by their primary research topic. This dataset is designed for tasks such as document classification, optical character recognition (OCR) testing, and multimodal document understanding.
Curated by: tegridy
Language: English
Format: PDF
Repo Structure
The dataset contains PDF files and their associated topic labels.… See the full description on the dataset page: https://huggingface.co/datasets/itstheprakash/research-papers.research_papersotu-taxa-paper-artifacts
OTU-Taxa paper artifacts
This repository contains stable processed artifacts used by the OTU-Taxa paper
and the frozen contracts needed to reproduce them.
Related repositories
Source code:
https://github.com/bio-ontology-research-group/otu-taxa-autoregressive
Comparator pipelines:
https://github.com/bio-ontology-research-group/microbiome_foundation_model_benchmarks
Pretrained OTU-Taxa model:
bio-ontology-research-group/otu-taxa
Model weights and reusable… See the full description on the dataset page: https://huggingface.co/datasets/bio-ontology-research-group/otu-taxa-paper-artifacts.filkom-research-papers
About Dataset
This dataset is a collection of FILKOM UB lecturers' research papers, collected from Google Scholar. It includes various fields such as title, abstract, authors, publication year, and more. The dataset is intended for research purposes and can be used to analyze lecturers' research topics, publication trends, and collaboration patterns.
Data Collection Methodology
The data was collected using web scraping techniques from Google Scholar. The scraping process… See the full description on the dataset page: https://huggingface.co/datasets/Lab-IS/filkom-research-papers.
