datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
minimax-m3-deepsearchqa-skill-eval
MiniMax M3 DeepSearchQA Skill Eval
Evaluates minimax/minimax-m3 on google/deepsearchqa using a Pi agent, You.com MCP tools, and a research skill optimized for this harness, model, and tool surface.
MiniMax M3 Medium Reasoning with the You.com research skill reached 74.85% adjusted F1 on DeepSearchQA, above the paper's GPT-5 High Reasoning F1 result. Public artifacts are available for inspection and reproduction.
Links
GitHub:… See the full description on the dataset page: https://huggingface.co/datasets/youdotcom/minimax-m3-deepsearchqa-skill-eval.Omni-DeepSearch
✨ Focus on Multimodal Audio Deep‑Search: Automated generation, filtering, and evaluation of multi‑hop reasoning benchmarks ✨
| 🧩 Fully Automated Pipeline | 🎵 Rich Audio Domains | 🧠 Multi‑Hop Reasoning QA | 🤖 Agentic Evaluation |
Omni-DeepSearch Benchmark
🎧 Multimodal Audio Focus – Designed for audio‑centric deep‑search tasks covering speech, music, bio‑acoustics, and environmental sounds.
🔄 Fully Automated Pipeline – End‑to‑end generation, multi‑stage… See the full description on the dataset page: https://huggingface.co/datasets/Kirito-Lab/Omni-DeepSearch.deepsearch-llama-finetune
DeepSearch Llama Finetune Dataset
Overview
The DeepSearch Llama Finetune Dataset is a specialized collection of high-quality, real-world prompts and responses, meticulously crafted for fine-tuning Llama-based conversational AI models. This dataset is optimized for:
Creativity: Responses are original, engaging, and leverage creative formats (Markdown, tables, outlines, etc.).
Effectiveness: Answers are highly relevant, actionable, and tailored for real-world applications.… See the full description on the dataset page: https://huggingface.co/datasets/enosislabs/deepsearch-llama-finetune.deepsearch-mini-shareGPT
DeepSearch Mini ShareGPT Dataset
Overview
The DeepSearch Mini ShareGPT Dataset is a curated collection of diverse, real-world prompts and highly effective responses, designed specifically for training and fine-tuning conversational AI models. This dataset emphasizes:
Efficiency: Answers are direct and to the point, maximizing information density.
Clarity: Explanations are easy to understand, even for complex topics.
Creativity: Responses are engaging, original, and often… See the full description on the dataset page: https://huggingface.co/datasets/enosislabs/deepsearch-mini-shareGPT.Deepsearch-test
