datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
marl-gpt-datasets
MARL-GPT Datasets
Offline expert trajectories from “MARL-GPT: Foundation Model for Multi-Agent Reinforcement Learning”.
Environments
This dataset includes trajectories from the three evaluation domains used in MARL-GPT: SMACv2 (StarCraft multi-agent combat), Google Research Football (GRF), and POGEMA (partially observable multi-agent pathfinding on grids).
Format
Trajectories are stored sequentially (no shuffling). Use the done flag to split the stream into… See the full description on the dataset page: https://huggingface.co/datasets/nortem/marl-gpt-datasets.sqp-tts-en
SQP TTS (English)
Synthesized speech for SQPsychConv_qwen-2.5, a synthetic CBT therapist-client
dialogue dataset (English).
Each configuration below corresponds to one TTS model. Load a single model
with:
from datasets import load_dataset
ds = load_dataset("sinselm/sqp-tts-en", "qwen3-tts")
Models included
qwen3-tts: https://huggingface.co/Qwen/Qwen3-TTS-12Hz-0.6B-Base
cosyvoice: https://huggingface.co/FunAudioLLM/Fun-CosyVoice3-0.5B-2512
fishaudio:… See the full description on the dataset page: https://huggingface.co/datasets/marleen-snsl/sqp-tts-en.MARL-GPT-LMAPFmoral-number-corpus
A Perspectivist Corpus of Numbers in Social Judgements
This is the dataset for A Perspectivist Corpus of Numbers in Social Judgements.
We constructed a corpus of moral and social judgements (questions are derived from the Commonsense Norm Bank) that asks people to fill in number ranges that do not change a given judgement.
Our corpus was crowdsourced from 30 annotators and contains 898 statements for a total of 3k annotations.
This work adds to available moral and social judgement… See the full description on the dataset page: https://huggingface.co/datasets/Marlon154/moral-number-corpus.Preference-Conditioned-Heterogeneous-MARL-Microgrid
SEGAN OPSD-Derived Microgrid Multiyear Benchmark
This repository contains the processed multiyear microgrid benchmark used for the study
“Preference-Conditioned Heterogeneous Multi-Agent Reinforcement Learning for Safe Microgrid Energy Management.”
Files
microgrid_opsd_multiyear.csv — processed hourly benchmark data.
opsd_multiyear_metadata.json — provenance, selected OPSD nodes, source-column mapping, scaling notes, and processing metadata.… See the full description on the dataset page: https://huggingface.co/datasets/Tristanchou/Preference-Conditioned-Heterogeneous-MARL-Microgrid.marl-world-model-lerobotstarai02This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "starai",
"total_episodes": 101,
"total_frames": 44639,
"total_tasks": 1,
"total_videos": 202,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:101"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Marlboro1998/starai02.annomi-tts-en
AnnoMI TTS (English)
Synthesized speech for the AnnoMI motivational interviewing dialogues (English).
Each configuration below corresponds to one TTS model. Load a single model
with:
from datasets import load_dataset
ds = load_dataset("sinselm/annomi-tts-en", "qwen3-tts")
Models included
qwen3-tts: https://huggingface.co/Qwen/Qwen3-TTS-12Hz-0.6B-Base
cosyvoice: https://huggingface.co/FunAudioLLM/Fun-CosyVoice3-0.5B-2512
fishaudio:… See the full description on the dataset page: https://huggingface.co/datasets/marleen-snsl/annomi-tts-en.dr-marl-papers
Distributionally-Robust RL & Cooperative MARL — top-tier paper index
A hand-reviewed index of 1,989 papers on distributionally-robust reinforcement learning
and cooperative multi-agent RL, drawn from a complete harvest of 75,819 accepted papers
at ICML, NeurIPS, ICLR, AAMAS and AISTATS (2010–2026).
Every candidate that survived the keyword filter was read and labelled one by one — 2,979
papers — rather than accepting an automatic classifier's output.
Buckets… See the full description on the dataset page: https://huggingface.co/datasets/Ngseo/dr-marl-papers.Futures_202306_202312Future dada on FG and sc.
historic-trading-data
