datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
papers-audio
Papers, as Audio
Narrated long-form lectures on statistics and machine learning, streamed by the
Papers, as Audio phone app.
library.json — the catalog the app reads (titles, durations, chapter markers, file paths)
audio/<id>.mp3 — one file per lecture (Kokoro-82M af_heart, 96 kbps mono)
lectures/<id>.md — transcripts
Published automatically from https://github.com/Arjun10g/Papers_Audio.
blue-arxiv-papersPaper2Video
Paper2Video: Automatic Video Generation From Scientifuc Papers
📃Arxiv | 🌐 Project Page | 💻Github
Dataset Description
The Paper2Video Benchmark includes 101 curated paper–video pairs spanning diverse research topics. Each paper averages about 13.3K words, 44.7 figures, and 28.7 pages, providing rich multimodal long-document inputs. Presentations contain on average 16 slides and run for about 6 minutes 15 seconds, with some reaching up to 14 minutes. Rather than… See the full description on the dataset page: https://huggingface.co/datasets/ZaynZhu/Paper2Video.nchlt_paper_sampleanv_paper_sample
