dblind/speakercard-1m-anonymous
SpeakerCard-1M (VoxCeleb release) Evidence-grounded Speaker Card corpus for in-the-wild speaker verification. SpeakerCard-1M is a speaker-centric resource built on VoxCeleb1/2 under a tool-first, LLM-last pipeline: ten acoustic probes extract field-level evidence, a schema separates relatively stable traits (gender, age band, accent, pitch band, timbre, language) from utterance-level states (emotion, channel, environment, speaking rate), and a constrained LLM verbalizes the… See the full description on the dataset page: https://huggingface.co/datasets/dblind/speakercard-1m-anonymous.
05
