datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PulseLM
PulseLM: A Foundation Dataset and Benchmark for PPG-Text Learning
Usage
from datasets import load_dataset, get_dataset_config_names, concatenate_datasets # datasets==4.5.0
dataset_names = get_dataset_config_names("Manhph2211/PulseLM")
print(f"Available datasets: {dataset_names}")
train_splits = [
load_dataset("Manhph2211/PulseLM", name, split="train").select_columns(["signal", "text", "qa"])
for name in dataset_names
]
combined =… See the full description on the dataset page: https://huggingface.co/datasets/Manhph2211/PulseLM.pulseedge-trading-ai-v5
PulseEdge Trading AI v5 — categorized lake
Private snapshot of the Hetzner storage box, split by data type.
v4 put packed tables at the repo root and everything else under a flat archives/ dump.
v5 keeps the same research lake but files live in typed folders.
Research use only. Not investment advice. Equities/ETFs only. No crypto.
Categories
Folder
What is in it
ohlcv/packed/
Combined Yahoo bars (1d/1h/15m/5m/1m) plus train/holdout splits
ohlcv/yahoo/… See the full description on the dataset page: https://huggingface.co/datasets/NickTheCoolst/pulseedge-trading-ai-v5.pulseedge-trading-ai-v2
PulseEdge Trading AI v2
Broader follow-up to NickTheCoolst/pulseedge-trading-ai (v1).
Equities/ETFs only. No crypto. No raw ITCH. No SEC archives.
Research use only. Not investment advice. Survivorship bias in the liquidity universe.
Use purged walk-forward splits. Label columns leak the future by construction.
What is new vs v1
v1
v2
Liquidity bar
$250k, 2y, $2
$80k, 1y, $1
Daily features (filtered)
24.1M / 6,896
25.6M / 7,505
Daily features (full… See the full description on the dataset page: https://huggingface.co/datasets/NickTheCoolst/pulseedge-trading-ai-v2.x402-market-pulse
x402 Market Pulse
A small, growing time series of how much USDC actually settles through x402 pay-per-call APIs on Base,
measured from on-chain Transfer logs rather than from provider-reported call counts. One row per
measurement; every row is a trailing 24-hour window ending at measured_at_utc.
Status: bootstrapping. 75 rows spanning 2026-09-22T13:34Z to 2026-09-26T07:00Z —
less than one independent day of data so far. Rows are kept to at most one per clock hour and a new one… See the full description on the dataset page: https://huggingface.co/datasets/nimapro1381/x402-market-pulse.PULSE
PULSE
A Synchronized Five-Modality Dataset for Multi-Modal Daily Activity Understanding
MoCap · EMG · Eye Tracking · IMU · Fingertip Pressure — all hardware-synced at 100 Hz
NeurIPS 2026 Evaluations & Datasets · under double-blind review · CC BY-NC 4.0
At a glance
40
9
5
7,789
volunteers
scenarios (S1–S8 + S9 motion primitives)
modalities @ 100 Hz
dense action segments
337 total recordings(304 task + 33 S9)
~9.7 h total (S1–S8: ~7.0 h)
<10 ms… See the full description on the dataset page: https://huggingface.co/datasets/velvet-pine-22/PULSE.open-pulse-hackathon-data-analysis
LauzHack Projects Dataset
Dataset Summary
This dataset contains comprehensive information about projects submitted to
LauzHack (EPFL's student-run hackathon) from 2023 to 2025. Each project
includes details about the project title, description, team members, awards, and
categories.
LauzHack is an annual 24-hour hackathon hosted at EPFL (École Polytechnique
Fédérale de Lausanne) in Lausanne, Switzerland, bringing together students and
hackers to create innovative solutions… See the full description on the dataset page: https://huggingface.co/datasets/SDSC/open-pulse-hackathon-data-analysis.worldcup-pulse-data
WorldCup Pulse Data
This Hugging Face Dataset repository is the single source of truth for WorldCup Pulse Lakehouse data. The GitHub code repo never commits generated data; GitHub Actions uploads this lakehouse output to this dataset repo.
Lakehouse layout
bronze/*.parquet # normalized raw API snapshots
silver/*.parquet # cleaned canonical entities and events
gold/*.parquet # dashboard-ready marts
state/last_run.json # incremental state
logs/*.csv… See the full description on the dataset page: https://huggingface.co/datasets/n2d/worldcup-pulse-data.PulseLM
PulseLM: A Foundation Dataset and Benchmark for PPG-Text Learning
Hung Manh Pham*
Jinyang Wu*
Xiao Ma
Yiming Zhang
Yixin Xu
Aaqib Saeed
Bin Zhu†
Zhou Pan†
Dong Ma†
* Equal contribution † Corresponding authors
Introduction
PulseLM is a multimodal framework that integrates PPG (Photoplethysmography) signal encoders with large language models for physiological signal understanding research. The project includes a large-scale… See the full description on the dataset page: https://huggingface.co/datasets/Ronilos/PulseLM.full-dubai-pulse120-physicsdojo137-pulseox-test4This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101_follower",
"total_episodes": 20,
"total_frames": 19763,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:20"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/LeRobot-worldwide-hackathon/120-physicsdojo137-pulseox-test4.quantum-control-pulse-instability-v0.1
quantum-control-pulse-instability-v0.1
What this dataset does
This dataset evaluates whether models can detect instability in quantum control pulse regimes.
Each row represents a simplified control scenario where quantum gates are implemented through microwave or optical pulse sequences.
The task is to determine whether the pulse regime remains stable or becomes unstable due to drift, noise, or synchronization failures.
Core stability idea
Quantum control… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/quantum-control-pulse-instability-v0.1.infer-pulse-eval
Infer Pulse Static Analysis Evaluation Dataset
Dataset Description
This dataset contains 523 C functions extracted from Meta's Infer static analyzer test suite, specifically the Pulse analyzer tests. It's designed for evaluating Large Language Models (LLMs) on static analysis tasks, particularly memory safety bug detection in C code.
Note: This is an evaluation-only dataset. All examples are provided in the test split.
Key Features
523 individual C… See the full description on the dataset page: https://huggingface.co/datasets/shubhamugare/infer-pulse-eval.race-pulse-parquetPulseBench-Select
PulseBench-Select
A benchmark for selected-option detection in document images.
PulseBench-Select contains 485 cleaned document images with ground-truth annotations for checkboxes, radio buttons, and marked answer choices. Each sample pairs a document image with public ground truth for the visible options that are selected.
Scoring methodology (GitHub): https://github.com/Pulse-Software-Corp/PulseBench-Select
Quick Start
from datasets import load_dataset
import… See the full description on the dataset page: https://huggingface.co/datasets/pulse-ai/PulseBench-Select.PulseFeed_Validated_datasethousehold-pulse-survey-hps-covid-19-vaccination-am
Household Pulse Survey (HPS): COVID-19 Vaccination among People with Disabilities
Description
Household Pulse Survey (HPS): HPS is a rapid-response survey of adults ages ≥18 years led by the U.S. Census Bureau, in partnership with seven other federal statistical agencies, to measure household experiences during the COVID-19 pandemic. Detailed information on probability sampling using the U.S. Census Bureau’s Master Address File, questionnaires, response rates, and bias… See the full description on the dataset page: https://huggingface.co/datasets/HHS-Official/household-pulse-survey-hps-covid-19-vaccination-am.PulseFeed_Data_Setpulse_validated_ticks_testPulse
