datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pimm_nasdaq_earningscall
Dataset Description (Categorized)
To support the development and evaluation of the Physics-Informed Acoustic Model (PIAM), we construct a multimodal dataset categorized as follows:
1. Sound Earnings Call Data
Number of Companies: 283 NASDAQ-listed corporations
Number of Recordings: 1,795 earnings call sessions
Total Audio Duration: Approximately 1,780 hours
Time Range: January 22, 2021 – June 29, 2025
Speaker Roles Labeled: CEO, CFO, CXO (and others where… See the full description on the dataset page: https://huggingface.co/datasets/soundai2016/pimm_nasdaq_earningscall.Onielkiryl-staselka-pimki-krystsina-drobysh
Пімкі
Metadata
Author: Кірыл Стаселька
Title: Пімкі
Narrator: Крысціна Дробыш
Source Group: Дзіцячыя
Source: https://knizhnyvoz.by/
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum split… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/kiryl-staselka-pimki-krystsina-drobysh.
