CoolFace
Datasetpublicgated

kapturecx/bolAIndia

bolAIndia Human-side speech from production call recordings, cut into utterance-level chunks by a two-engine VAD (Silero + TEN) and transcribed by third-party ASR providers. Each row keeps the transcript, the provider's confidence, and full provenance back to the source recording. Sources One config per transcription system, so their output stays separable. config (source_id) provider model hours rows shards vendor-a vendor-a undisclosed 420.03 480774… See the full description on the dataset page: https://huggingface.co/datasets/kapturecx/bolAIndia.

sourceHugging Faceupdated 10m agoView on Hugging Face
1likes12kdownloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.

kapturecx/bolAIndia · CoolFace