datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sft-robo2-data-place_phone_stand
SFT-Robo2 Expert Data: place_phone_stand
Expert demonstration dataset for the place_phone_stand task from RoboTwin 2.0, collected for SFT training of OpenVLA-OFT following the SimpleVLA-RL paper (arXiv:2509.09674).
Dataset Structure
aloha/ # Raw ALOHA-format HDF5 (94GB)
train/ # 950 episodes
episode_0.hdf5
...
val/ # 50 episodes
episode_0.hdf5
...
rlds/ # RLDS/TFDS format (7.8GB)
1.0.0/… See the full description on the dataset page: https://huggingface.co/datasets/Louisnguyen/sft-robo2-data-place_phone_stand.phone-detection-dataphone_left_to_right_red_up_slot_1280x720_data1_51epPhone_Timings_Database
📖 TajweedAI: Quranic Phoneme Timing Benchmark (Phases 1, 2 & 3)
📌 Project Overview
TajweedAI evaluates Quranic recitation accuracy by analyzing both pronunciation (phoneme classification) and timing (rule duration evaluation).
This benchmark provides empirical, tempo-normalized duration boundaries for all 70 Quranic phonemes derived from forced alignments (MFA trained on Quranic audio) across 7 master reference reciters:
Sheikh Mahmoud Khalil Al-Husary (Gold… See the full description on the dataset page: https://huggingface.co/datasets/AhmedTamertechno1/Phone_Timings_Database.Mobile_Phone_Dataset_Smartphone_and_Feature_Phone
Mobile Phone Dataset — Smartphone and Feature Phone (Sample)
⚠️ This is a free sample subset for evaluation purposes only.The full dataset (3,000+ HD images) is available for commercial licensing.Contact: sales@datacluster.ai · datacluster.ai
Dataset Summary
This dataset is an extremely challenging collection of original mobile phone images, crowdsourced from over 1,000 urban and rural areas. Every image is manually reviewed and verified by computer vision… See the full description on the dataset page: https://huggingface.co/datasets/Dataclusterlabspvtltd/Mobile_Phone_Dataset_Smartphone_and_Feature_Phone.phone-scam-datasetphone-and-webcam-dataset
Video Dataset - 1,300+ files
The dataset comprises 1,300+ videos of 300+ people captured using mobile phones (including Android devices and iPhone) and webcams under varying lighting conditions. It is designed for research in face detection, object recognition, and event detection, leveraging high-quality videos from smartphone cameras and webcam streams. — Get the data
Dataset characteristics:
Characteristic
Data
Description
Each person recorded 4 videos… See the full description on the dataset page: https://huggingface.co/datasets/ud-biometrics/phone-and-webcam-dataset.Phone-FA-EN-AR-Dataset
بسم الله
این مخزن شامل دو دادگان فارسی-انگلیسی-عربی و عربی-انگلیسی است که به کمک یک موتور ای-اسپیک تغییر یافته جمع آوری شده است.
(شما می توانید این برنامه را در اینجا و اینجا پیدا کنید)
هم چنین یک رابط کاربری برای جمع آوری دادگان صوتی در اینجا قرار داده شده است.
دادگان اول (فارسی-انگلیسی-عربی)
دادگان اول (دادگان فارسی-انگلیسی-عربی) خوانش تاک-بک یک گوشی اندرویدی(گالکسی اس۱۰ - اندروید ۹) است.
این دیتا ست شامل موارد زیر است:
حروف کیبرد فارسی و انگلیسی
تمام… See the full description on the dataset page: https://huggingface.co/datasets/mah92/Phone-FA-EN-AR-Dataset.Taiwan_Mandarin_Speech_Data_by_Mobile_Phone_Reading
Dataset Card for Nexdata/Taiwan_Mandarin_Speech_Data_by_Mobile_Phone_Reading
Dataset Summary
This dataset is just a sample of Taiwan Mandarin Speech dataset(paid dataset) by mobile phone reading.The data collects 204 Taiwan residents with 450 sentences for each speaker. The recorded is rich in content, including economy, entertainment, news, spoken language, numbers, letters, etc., covering general scenes and human-computer interaction scenes. Manual transcription of text… See the full description on the dataset page: https://huggingface.co/datasets/lianghsun/Taiwan_Mandarin_Speech_Data_by_Mobile_Phone_Reading.so101-phone-teleop-dataThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so100_follower",
"total_episodes": 1,
"total_frames": 383,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 500,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/hproc/so101-phone-teleop-data.Kawthar-AR_EN-Public-Phone-Audio-Dataset
بسم الله
This dataset text data is derived from here.
Audio files are gathered by the help of Arabic team: Planet Blind Tech (PBt).
Thank you Shams Eddin (from Algeria).
Ayoub-AR_EN-Public-Phone-Audio-Dataset
بسم الله
This dataset text data is derived from here.
Audio files are gathered by the help of Arabic team: Planet Blind Tech (PBt).
Thank you Shams Eddin (from Algeria).
phone-asr-dataphone-to-so100-dataset1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"ee.x",
"ee.y",
"ee.z",
"ee.wx",
"ee.wy",
"ee.wz",
"ee.gripper_pos"
],
"shape": [
7
]
}… See the full description on the dataset page: https://huggingface.co/datasets/marshin68/phone-to-so100-dataset1.Khadijah-FA_EN-Public-Phone-Audio-Dataset
Text got from here.
All chinese letters be replaced with "chinese letter" because espeak reads them so...
Remove all persian/english single alphabet as they are not read correctly(same as espeak) by my reader...
Replace these chars with space, as they where not read correctly(same as espeak) :
🔺
@
/
)
(
]
[
▪️
🔹️
🔷
🔶
🔆
📌
⚫️
™
❤
🏆
◉
👍
🔥
😱
👌
📍
✈︎
☁︎
⚡️
➖
🍅
😁
👇
🤩
😢
🥰
😁
🤯
🤲
👏
🎬
✊
💙
🤝
😮
😎
😇
🙏
🥀
⬇️⬇️
🌀
🖤
😵
🍿
👇🏼
🤔
🎉
🥰
✅
🆔
😍
🤣
🔴
🪐
🕊
🗓
🇺🇳
✴️
🔹️… See the full description on the dataset page: https://huggingface.co/datasets/mah92/Khadijah-FA_EN-Public-Phone-Audio-Dataset.phone-to-so100-datasetThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"ee.x",
"ee.y",
"ee.z",
"ee.wx",
"ee.wy",
"ee.wz",
"ee.gripper_pos"
],
"shape": [
7
]
}… See the full description on the dataset page: https://huggingface.co/datasets/marshin68/phone-to-so100-dataset.Musa-FA_EN-Public-Phone-Audio-DatasetSame as: https://huggingface.co/datasets/mah92/Khadijah-FA_EN-Public-Phone-Audio-Dataset
phone_left_to_right_red_up_slot_1280x720_data_71epMobile_Phone_Evolution_DatasetMotahare-FA_EN_AR-Public-Phone-Audio-DatasetUnique007_french_phone_1421_samples_16khzCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données Unique007/french_phone_1421_samples_16khz.
nigeria-phone-sales-datasetdata_phoneNexdata_French_Conversational_Speech_Data_by_Mobile_PhoneCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données Nexdata/French_Conversational_Speech_Data_by_Mobile_Phone.
Nexdata_French_Speech_Data_by_Mobile_Phone_ReadingCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données Nexdata/French_Speech_Data_by_Mobile_Phone_Reading.
Unique007_french_phone_16khzCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données Unique007/french_phone_16khz.
Motahare-AR_EN-Public-Phone-Audio-DatasetNexdata_French_Speech_Data_by_Mobile_Phone_GuidingCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données Nexdata/French_Speech_Data_by_Mobile_Phone_Guiding.
Hassan-AR_EN-Public-Phone-Audio-DatasetNexdata_French_Speech_Data_by_Mobile_PhoneCe répertoire est vide, il a été créé pour améliorer le référencement du jeu de données Nexdata/French_Speech_Data_by_Mobile_Phone.
