CoolFace
Datasetpublic

sjhuskey/enenlhet-wav2vec2-dataset

Enenlhet Wav2Vec2 Dataset This dataset contains preprocessed audio features and tokenized text for training Wav2Vec2 models on the Enenlhet language. Dataset Summary Train: 3,053 examples Test: 170 examples Validation: 170 examples Total: 3,393 examples Features input_values: Preprocessed audio features (16kHz, normalized float32 arrays) labels: Tokenized text as integer sequences Usage from datasets import load_dataset # Load… See the full description on the dataset page: https://huggingface.co/datasets/sjhuskey/enenlhet-wav2vec2-dataset.

sourceHugging Facecc-by-nc-4.0updated 1y agoView on Hugging Face
0likes37downloads
settings

This repository belongs to sjhuskey on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameenenlhet-wav2vec2-dataset
visibilitypublic
licencecc-by-nc-4.0
gatedno
ownersjhuskey
Account settings
sjhuskey/enenlhet-wav2vec2-dataset · CoolFace