datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
how2sign-asl-landmarks
How2Sign ASL Landmarks
Sentence-level MediaPipe landmark cache for the How2Sign ASL translation dataset. This dataset stores preprocessed landmark arrays only, not MP4 clips.
The rows come from the local How2Sign realigned CSV files and front-view raw videos. Each successful row is one sentence segment sampled at a 25 FPS cap and stored in compressed NPZ shards.
Why This Dataset Exists
The public How2Sign data has known alignment/completeness issues: some text rows do not… See the full description on the dataset page: https://huggingface.co/datasets/martinctl/how2sign-asl-landmarks.google_landmarks_places
Google Landmarks places
Google Landmarks is a great dataset, but it lacks geospatial information about the places. This dataset fills
this gap by providing latitude and longitude for each landmark. The dataset also contains the name of the landmark from OpenStreetMap
and information about the country, the province/state, and the city/village where the landmark is located. This information was collected from OSM
via Nominatim.
Optimized_Video_Facial_Landmarks
Dataset Card for 478-Point Normalized 3D Facial Landmark Dataset
Dataset Description
This dataset provides pre-extracted, normalized 3D facial landmark features derived from the Video Emotion dataset. It is optimized for efficient training of emotion recognition and facial analysis models, bypassing the need to process large raw video files.
License: The extracted feature data in this Parquet file is licensed under Apache 2.0. Note that the original source video files may… See the full description on the dataset page: https://huggingface.co/datasets/PSewmuthu/Optimized_Video_Facial_Landmarks.slovo-full-landmarks
Slovo landmarks base
Base landmarks dataset built from full Slovo videos with MediaPipe Tasks HolisticLandmarker.
Emotion_Video_Facial_Landmarks
Dataset Card for 478-Point Normalized 3D Facial Landmark Dataset
Dataset Description
This dataset provides pre-extracted, normalized 3D facial landmark features derived from the Video Emotion dataset. It is optimized for efficient training of emotion recognition and facial analysis models, bypassing the need to process large raw video files.
License: The extracted feature data in this CSV file is licensed under Apache 2.0. Note that the original source video files may… See the full description on the dataset page: https://huggingface.co/datasets/PSewmuthu/Emotion_Video_Facial_Landmarks.google_landmarks_photos
Dataset Card for "google_landmarks_photos"
More Information needed
Emotion_Video_Facial_Landmarks
Dataset Card for 478-Point Normalized 3D Facial Landmark Dataset
Dataset Description
This dataset provides pre-extracted, normalized 3D facial landmark features derived from the Video Emotion dataset. It is optimized for efficient training of emotion recognition and facial analysis models, bypassing the need to process large raw video files.
License: The extracted feature data in this CSV file is licensed under Apache 2.0. Note that the original source video files may… See the full description on the dataset page: https://huggingface.co/datasets/mac26/Emotion_Video_Facial_Landmarks.welsh-speech-landmarks
Welsh Speech Dataset - Facial Landmarks
68-point facial landmarks (ibug68 template) from the Welsh Speech Dataset.
Contents
Facial landmarks for every frame
68 3D points per frame (x, y, z coordinates)
Format: Parquet
Manual annotation using ibug68 template
Format
The landmarks.parquet file contains:
Column
Description
speaker_id
Speaker identifier (1-33)
phrase_id
Phrase identifier (1-10)
frame_id
Frame identifier (e.g., "001", "002")… See the full description on the dataset page: https://huggingface.co/datasets/arvinsingh/welsh-speech-landmarks.NSL-Alphabet-Landmarksinclude-landmarks
