CoolFace
Datasetpublic

michaelcacioli/Neapolitan-Spoken-Corpus

Neapolitan Spoken Corpus (NSC) A corpus of read Neapolitan speech for ASR evaluation, with a validated Neapolitan–Italian lexicon, LOSO fine-tuning splits, trained LoRA adapters, metric implementations, per-clip results, and error annotations. This release supersedes the earlier 141-clip single-speaker version of this repository. The earlier release corresponds to Speaker S1 of the present corpus; the old audioData/ and transcripts.csv are replaced by data/audio/ and… See the full description on the dataset page: https://huggingface.co/datasets/michaelcacioli/Neapolitan-Spoken-Corpus.

sourceHugging Facecc-by-nc-4.0updated 3mo agoView on Hugging Face
4likes308downloads
14 commits on main
a8825cc3mo ago

Remove superseded files: old transcripts.csv, requirments.txt, code/, and 141-clip audioData/ (replaced by data/audio and data/metadata.csv)

anonymous-nsc-author
fd6738d3mo ago

Full release: expanded NSC (591 clips), lexicon, LOSO splits, LoRA adapters, metrics, annotations, rebuttal analyses

anonymous-nsc-author
8e21d613mo ago

Expansion

anonymous-nsc-author
75a25791y ago

Delete audioData/test

anonymous-nsc-author
5e829a21y ago

Upload 141 files

anonymous-nsc-author
4458f4c1y ago

Create audioData/test

anonymous-nsc-author
fe7740e1y ago

Create transcribe_whisper.py

anonymous-nsc-author
18645a81y ago

fix format mistake

anonymous-nsc-author
a2d699c1y ago

Create code/generate_json.py

anonymous-nsc-author
493ac741y ago

Create code/evaluate_metrics.py

anonymous-nsc-author
1b500431y ago

Create transcripts.csv

anonymous-nsc-author
849593b1y ago

Create requirments.txt

anonymous-nsc-author
c5a48981y ago

Update README.md

anonymous-nsc-author
b5211c41y ago

initial commit

anonymous-nsc-author