CoolFace
Datasetpublic

distil-whisper/ami-sdm

The AMI Meeting Corpus consists of 100 hours of meeting recordings. The recordings use a range of signals synchronized to a common timeline. These include close-talking and far-field microphones, individual and room-view video cameras, and output from a slide projector and an electronic whiteboard. During the meetings, the participants also have unsynchronized pens available to them that record what is written. The meetings were recorded in English using three different rooms with different acoustic properties, and include mostly non-native speakers. \n

sourceHugging Facecc-by-4.0updated 3y agoView on Hugging Face
1likes67downloads
16 commits on main
0fb55493y ago

fix license

sanchit-gandhi
146e87a3y ago

add readme

sanchit-gandhi
ae2cc143y ago

add optional gating

sanchit-gandhi
f3e11213y ago

add optional gating

sanchit-gandhi
a7d12693y ago

add optional gating

sanchit-gandhi
88801213y ago

add readme

sanchit-gandhi
6ec56973y ago

remove local transcriptions

sanchit-gandhi
f2993913y ago

fix bug in csv loading

sanchit-gandhi
5723b883y ago

ami boom

sanchit-gandhi
cd39f8d3y ago

uniform dataset ids

sanchit-gandhi
8b785d03y ago

up

sanchit-gandhi
6948e1f3y ago

consistency

Sanchit Gandhi
28aa47a3y ago

Saving transcriptions for split test

Sanchit Gandhi
c73e42e3y ago

Saving transcriptions for split validation

Sanchit Gandhi
a56bb6f3y ago

Saving transcriptions for split train

Sanchit Gandhi
975ded13y ago

initial commit

Sanchit Gandhi