datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
my-dataclaw-data
Coding Agent Conversation Logs
This is a performance art project. Anthropic built their models on the world's freely shared information, then introduced increasingly dystopian data policies to stop anyone else from doing the same with their data - pulling up the ladder behind them. DataClaw lets you throw the ladder back down. The dataset it produces is yours to share.
Exported with DataClaw.
Tag: dataclaw - Browse all DataClaw datasets
Stats
Metric
Value… See the full description on the dataset page: https://huggingface.co/datasets/peteromallet/my-dataclaw-data.my-dataclaw-data
Coding Agent Conversation Logs
This is a performance art project. Anthropic built their models on the world's freely shared information, then introduced increasingly dystopian data policies to stop anyone else from doing the same with their data - pulling up the ladder behind them. DataClaw lets you throw the ladder back down. The dataset it produces is yours to share.
Exported with DataClaw.
Tag: dataclaw - Browse all DataClaw datasets
Stats
Metric
Value… See the full description on the dataset page: https://huggingface.co/datasets/A99311/my-dataclaw-data.My_Dataset
IndicFCB — Agriculture (English dialogues)
303 English user-side dialogues for Indian government agricultural services,
classified along the IndicFCB benchmark axes. One row per (task x persona) rendering.
Columns
task_id — stable identifier
turns_axis — single_turn | multi_turn
steps_axis — single_step | multi_step
type_axis — simple | parallel_multiple | irrelevance
usecase — service+goal tag (e.g. pmkisan_income, kcc_credit)
intent — one-line summary of the… See the full description on the dataset page: https://huggingface.co/datasets/saiteja2311/My_Dataset.mydataset2
Amod/mental_health_counseling_conversations
This dataset is a compilation of high-quality, real one-on-one mental health counseling conversations between individuals and licensed professionals. Each exchange is structured as a clear question–answer pair, making it directly suitable for fine-tuning or instruction-tuning language models that need to handle sensitive, empathetic, and contextually aware dialogue.
Since its public release in 2023, it has been downloaded over 100,000… See the full description on the dataset page: https://huggingface.co/datasets/onyi666/mydataset2.
