datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
tool-reasoning-sft-RESEARCH-grill-lab-browsecomp-plus-runs-data-cleaned-rectified
Tool-Reasoning SFT — BrowseComp-Plus Runs (Cleaned & Rectified)
Multi-turn tool-use reasoning trajectories derived from grill-lab/browsecomp-plus-runs, converted to a structured SFT format following the interstellarninja/hermes_reasoning_tool_use convention.
Source
Based on the execution trajectories from "Revisiting Text Ranking in Deep Research" (arXiv:2602.21456):
Original data: grill-lab/browsecomp-plus-runs (MIT)
Format
Each row contains a messages… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-RESEARCH-grill-lab-browsecomp-plus-runs-data-cleaned-rectified.browsecomp-plus-glm52-fp8-trie-event-replay
GLM-5.2 FP8 BrowseComp-Plus Trie Event Replay
This manually gated dataset contains a captured BrowseComp-Plus agent workload
served by GLM-5.2 FP8 on SGLang with CPU BM25 retrieval and evaluation
concurrency eight: all 830 benchmark queries, one closed terminal trajectory
each, recorded as Trie schema-v2 causal event streams and directly replayable
against any OpenAI-compatible inference endpoint.
Quick start: download, cd, replay
hf download… See the full description on the dataset page: https://huggingface.co/datasets/weili-0234/browsecomp-plus-glm52-fp8-trie-event-replay.
