datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fiverr-gigs
Fiverr Gigs — community dataset
Public Fiverr listing metadata collected via the
fiverr-gig-optimizer
Claude Code skill. Grown via opt-in contributions from users who run
contribute.py. Licensed CC-BY-4.0.
What's in it
Each record follows the canonical schema (title, category/subcategory, tier
prices, delivery days, tags, rating, review_count, gig_count_in_search,
currency). Prices are normalized to USD.
Privacy
Contributions are anonymized before… See the full description on the dataset page: https://huggingface.co/datasets/Ahad690/fiverr-gigs.growthkit-trends
GrowthKit Trends
A community, opt-in, federated dataset of public, anonymized short-form
short-form-video trend and benchmark observations, contributed by users of the
open-source GrowthKit Claude
Code skill. It improves GrowthKit's default benchmarks over time so every
founder starts from better, source-tagged ranges instead of fabricated numbers.
Honesty first. GrowthKit never lets a model invent a market metric. Numbers
come from deterministic scripts run on a founder's own… See the full description on the dataset page: https://huggingface.co/datasets/Ahad690/growthkit-trends.5gnr-pusch-iq-dmrs
5G NR PUSCH IQ DMRS Captures
Frequency-domain PUSCH IQ captures from an OAI 5G NR gNB (NI USRP X410 / USRP B210) with per-device IMSI labels embedded directly in each capture record. Collected over-the-air with commercial Quectel RM520N-GL modems and a software-defined USRP B210 UE.
Dataset summary
Filtering: only captures with a strongly visible DMRS RE comb (active/quiet power ratio ≥ 1.5×) are accepted
Format: v4 binary (.bin) — see nr_pusch_capture_oai for… See the full description on the dataset page: https://huggingface.co/datasets/ahancock516/5gnr-pusch-iq-dmrs.chess-a6-to-d6This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 3588,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/chess-a6-to-d6.361-chessbot-chess_movementchess_a8_to_d8This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 3542,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/chess_a8_to_d8.chess-c5-to-b6This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 3588,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/chess-c5-to-b6.chess-b5-to-c7This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 3587,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/chess-b5-to-c7.chess-b5-to-b6This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 3589,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/chess-b5-to-b6.chess-d6-to-d5This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 3590,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/chess-d6-to-d5.sms_spamrack_test_tube_2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "bimanual_so",
"total_episodes": 10,
"total_frames": 5702,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/rack_test_tube_2.rack_test_tube_4This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 5769,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/rack_test_tube_4.rack_test_tube_3This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 5796,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/rack_test_tube_3.rack_test_tube_5This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 5751,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/rack_test_tube_5.rack_test_tube_top_view_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 5977,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/rack_test_tube_top_view_1.rack_test_tube_6This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101",
"total_episodes": 10,
"total_frames": 5793,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/rack_test_tube_6.rack_test_tube_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "bimanual_so",
"total_episodes": 10,
"total_frames": 5708,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:10"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahad-j/rack_test_tube_1.AHaBench
Usage Guideline
This dataset and accompanying code are released under the Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) License.You may copy, share, and adapt the materials for non-commercial research and educational purposes, provided that you:
Give proper credit to the authors.
Include a link to the license.
Clearly indicate any modifications.
👉 Commercial use is strictly prohibited without prior written consent from the authors.… See the full description on the dataset page: https://huggingface.co/datasets/anonymous9268/AHaBench.ahat_task_20k_0630AHaBench
Usage Guideline
This dataset and accompanying code are released under the Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) License.You may copy, share, and adapt the materials for non-commercial research and educational purposes, provided that you:
Give proper credit to the authors.
Include a link to the license.
Clearly indicate any modifications.
👉 Commercial use is strictly prohibited without prior written consent from the authors.… See the full description on the dataset page: https://huggingface.co/datasets/o0oMiNGo0o/AHaBench.so101_grab_boxThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 0,
"total_frames": 0,
"total_tasks": 0,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ahaozi/so101_grab_box.ahammadmejbah_financial-risk-classification
Financial Risk Classification
A synthetic dataset for supervised learning using Logistic Regression, KNN, and
Dataset Info
Source: Kaggle
Original Size: 0.55 MB
Kaggle Downloads: 184
Files: 1
Files
Financial Risk Classification Dataset.csv
Mirrored from Kaggle
ahat_task_40kGenerate by the GPT-4.1 model, there are 5 prompts to use for the diversity in plan length.
The data contains two parts.
5k data without feedback
50 scenes from hssd, sample 1 persona, 5 different prompts, 2 iterations, generate 10 task in one call.
35 data with feedback; but there are some bugs lead to no feedback in object_distribution and plan_length_distribution.
50 scenes from hssd, sample 2 persona, 5 different prompts, 7 iterations, generate 10 task in one call.
ahat_task_20kGenerate by the GPT-4.1 model, there are 5 prompts to use for the diversity in plan length.
50 scenes from hssd, sample 2 persona, 5 different prompts, 4 iterations, generate 10 task in one call.
Generate 20k data in total.
ahat_task_20k_task_generate_update_sceneafrica-ghana-trends-in-road-accidents-and-casualties-brong-ahafo-region-cf893603
Trends in Road Accidents and Casualties Brong Ahafo Region | Africa (Ghana Open Data)
293 rows - 1 Africa country/area - 1991-2011 - source table - Engineered by Electric Sheep Africa
TL;DR
This dataset contains 293 rows from Ghana Open Data, covering Trends in Road Accidents and Casualties Brong Ahafo Region. It is published as ML-ready Parquet with consistent Hugging Face metadata, source provenance, and analysis-friendly loading examples.
What… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-ghana-trends-in-road-accidents-and-casualties-brong-ahafo-region-cf893603.ahat_task_40k_task_generate_update_sceneWorldHappinessDataset2015-2023ahat_task_50k_0701
