datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
sample_result_of_claudeqwen9b-coop-claude-code
qwen9b-coop-claude-code
Two-agent cooperative coding trajectories generated by running
CooperBench in coop mode on
the CooperData task set, using
Qwen/Qwen3.5-9B as the model and Claude Code (claude_code) as the
agent framework. Each pair runs two agents in parallel — one per feature —
coordinating via Redis messaging and a shared git remote.
The matched solo (single-agent) baseline is at
CooperBench/qwen9b-solo-claude-code.
Same task corpus, same model, same agent — only the… See the full description on the dataset page: https://huggingface.co/datasets/CooperBench/qwen9b-coop-claude-code.qwen9b-solo-claude-code
qwen9b-solo-claude-code
Single-agent coding trajectories generated by running
CooperBench in solo mode on
the CooperData task set, using
Qwen/Qwen3.5-9B as the model and Claude Code (claude_code) as the
agent framework. One agent implements both features in each task.
The matched coop (two-agent) version is at
CooperBench/qwen9b-coop-claude-code.
Same task corpus, same model, same agent — only the coordination differs, so
together they isolate the cooperation deficit.
At a… See the full description on the dataset page: https://huggingface.co/datasets/CooperBench/qwen9b-solo-claude-code.multilingual-llm-jokes-4o-claude-gemini
Rapidata Generated Joke Preference Dataset
We collected 1'000'000+ human opinions on the jokes generated by state-of-the-art LLMs to decide which model is the funniest. The labelers are shown a joke in their language and asked to answer 'Yes' or 'No' to the question 'Is this joke funny?'.
It took us less than 5 days to get all of the responses.
The jokes are evenly distributed across 5 languages: English, Arabic, Japanese, Vietnamese, Portuguese and across 4 model… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/multilingual-llm-jokes-4o-claude-gemini.eleicoes-municipais-2024claude_opus-4.6_4.7_coding_reasoningqwen9b-coop-claude-code-compressed
qwen9b-coop-claude-code-compressed-ak
Synthetic compressed cooperative agent trajectories derived from
CooperBench/qwen9b-coop-claude-code.
Each raw pair (two LLM coding agents on overlapping features in the same
repo, communicating via Redis messaging + a team git remote) is condensed
into an idealized version: wasted steps dropped, broken submission rituals
fixed, missing cooperation events (coop-send/coop-broadcast/coop-recv,
git fetch/diff team) inserted where the real pair… See the full description on the dataset page: https://huggingface.co/datasets/CooperBench/qwen9b-coop-claude-code-compressed.european-countries
European Countries
A small reference dataset of European countries with capital, population,
area, currency, ISO codes, and EU membership status.
Population and area figures are approximate recent estimates. Russia and
Turkey are transcontinental; full territory figures are reported.
hsqa-claude
HSQA‑Claude Medical QA Dataset
A parallel dataset of technical and simplified answers to real-world consumer medical questions, generated using Claude 3.5 Sonnet and aligned for medical text simplification research.
🧠 Overview
Purpose: Support training and evaluation of models for medical text simplification, readability control, and audience‑tailored generation.
Source Questions: Based on 3,000+ consumer health queries from the HealthSearchQA dataset.
Answer Styles:… See the full description on the dataset page: https://huggingface.co/datasets/DNivalis/hsqa-claude.Claude-sonnetclaude-reviewed-sport-sustainability-papers
Claude Reviewed Sport Sustainability Papers
Dataset description
This dataset encompasses 16 papers (some of them divided in several parts) related to the sustainability of sports products and brands.
This information lies at the core of our application and will come into more intesive use with the next release of GreenFit AI.
The papers were analysed with Claude 3.5 Sonnet (claude-3-5-sonnet-20241022) and they were translated into:
Their title (or a Claude-inferred… See the full description on the dataset page: https://huggingface.co/datasets/greenfit-ai/claude-reviewed-sport-sustainability-papers.emotion_audio_dataset_preprocess_claudered_team_agent_analysis_claude_reconstruction_test_results
red_team_agent_analysis_claude_reconstruction_test_results
This dataset was automatically uploaded from the red-team-agent repository.
Dataset Information
Original file: claude_reconstruction_test_results.csv
Source path: /home/ubuntu/red-team-agent/red_team_agent/analysis/claude_reconstruction_test_results.csv
Validation: Valid CSV with 100 rows, 9 columns (0.7MB)
Usage
import pandas as pd
from datasets import load_dataset
# Load using datasets library… See the full description on the dataset page: https://huggingface.co/datasets/aq1048576/red_team_agent_analysis_claude_reconstruction_test_results.
