datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mcp-agent-trajectory-benchmark
MCP Agent Trajectory Benchmark
A benchmark dataset of 49 MCP (Model Context Protocol) agent trajectories (38 single-pass + 11 multi-conv) with complete tool-use traces in the ATIF v1.2 (Agent Trajectory Interchange Format) format. Each agent operates in a distinct business domain with custom tools, realistic user conversations, and full execution traces.
Designed for training and evaluating tool-use / function-calling capabilities of LLMs.
Overview
Item
Details… See the full description on the dataset page: https://huggingface.co/datasets/obaydata/mcp-agent-trajectory-benchmark.Pantheon-Agent-Trajectory
🏛️ Pantheon Agent Trajectory Gallery
Curated end-to-end agent runs from PantheonOS — an open multi-agent framework for scientific computing.
Each "trajectory" captures a complete chat session: the user prompt, every reasoning/tool step the agent(s) took, the code that was run, the figures that were produced, and the final report. Trajectories are fully inspectable and reproducible, designed for transparency, teaching, and benchmarking.
🔗 Browse the gallery (live):… See the full description on the dataset page: https://huggingface.co/datasets/NaNg/Pantheon-Agent-Trajectory.A-Survey-for-LLM-Agent-Trajectory-Analysis
A Survey for LLM Agent Trajectory Analysis
This dataset repository hosts the survey paper A Survey for LLM Agent Trajectory Analysis: From Failure Attribution to Enhancement and a structured metadata snapshot of the companion paper collection from Awesome-LLM-Agent-Trajectory-Analysis.
The repository is intended for discovery, citation, and lightweight analysis of the literature around LLM agent trajectory analysis, including failure attribution, trajectory-based debugging… See the full description on the dataset page: https://huggingface.co/datasets/RobinChen2001/A-Survey-for-LLM-Agent-Trajectory-Analysis.mcp-agent-trajectory-benchmark
⚡ Model Context Protocol (MCP) & Advanced Tool‑Use Alignment Tiers
Official Enterprise Data Repository by springofwindslabs
👉 Looking for full production data?
The complete 1,000-row standard volume and 2,300+ row mutually exclusive, non-overlapping extended package are fully available for commercial deployment via our official procurement gateway:
➔… See the full description on the dataset page: https://huggingface.co/datasets/springofwindslabs/mcp-agent-trajectory-benchmark.Agent-Trajectory-Dataset
Description
본 데이터셋은 심층 검색, 데이터 분석, 산업 리서치 등 사무 환경에서 수행되는 다양한 작업 시나리오를 포함하며, 완전한 멀티턴 추론 과정과 도구 호출 체인으로 구성되어 있습니다. 에이전트의 계획 수립 능력 분석, 도구 선택 전략 연구 및 작업 품질 평가를 지원하도록 설계되었으며, 에이전트 학습 및 평가를 위한 구조화된 벤치마크로 활용할 수 있습니다.
자세한 내용은 아래 링크를 참고해 주세요: https://ko.nexdata.ai/datasets/llm/2185?source=hf.kr
Specifications
Data content
OpenClaw를 통해 생성된 에이전트 트래젝토리 데이터
Category
심층 검색, 데이터 분석, 산업 리서치
Data volume
5,300
Model… See the full description on the dataset page: https://huggingface.co/datasets/Nexdata-kr/Agent-Trajectory-Dataset.Agent-Trajectory-Data-Sample
Agent-Trajectory-Dataset
Description
This dataset covers office-based scenarios such as in-depth searches, data analysis, and industry research, encompassing complete multi-turn reasoning trajectories and tool-calling chains. It is designed to support the analysis of agent planning capabilities, research into tool selection strategies, and quality assessment, providing a structured benchmark for agent training and evaluation.
For more details, please refer to the… See the full description on the dataset page: https://huggingface.co/datasets/Nexdata-AI/Agent-Trajectory-Data-Sample.agent_trajectory_reviewsAgent-Trajectory-2.8kembodied_web_agent_outdoor_trajectory
