datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
amazon_massive_intent_en-USResearch-Intent-Judge
Research Intent — LLM-as-Judge
▶️ Watch the Video
LLM-as-Judge annotations for research paper intent classification, collected
through the Echo-DSRN collaborative platform during the OpenAIRE AI Hackathon 2026.
The dataset has one split per judge model (Gemma_4_E4B_it_GGUF,
Qwen3.6_35B_A3B_GGUF, Bonsai_8B_gguf, ...) plus a human_annotations
split with curator annotations. Split names use underscores in place of the
dashes in model names (HF does not allow dashes in split… See the full description on the dataset page: https://huggingface.co/datasets/ethicalabs/Research-Intent-Judge.amazon_massive_intent_zh-CNamazon_massive_intent_de-DEforge-intentdata
Forge Intent Dataset
Version: 1.0.0
amazon_massive_intent_am-ETamazon_massive_intent_hi-INamazon_massive_intent_sw-KEamazon_massive_intent_ar-SAamazon_massive_intent_ja-JPintents-for-eval
Purpose. This dataset was collected specifically for intent-parser benchmarking, independently from any OVOS skill. Skill-derived utterances tend to overfit the exact phrasings a plugin was tuned on; this data is drawn from a disjoint source so it measures whether an OVOS intent plugin generalizes rather than memorizes. It is part of the OVOS intent-classification datasets used by the OVOS Plugin Arena intent benchmark.
Funding
Developed by TigreGotico for OpenVoiceOS as part… See the full description on the dataset page: https://huggingface.co/datasets/OpenVoiceOS/intents-for-eval.amazon_massive_intent_th-THamazon_massive_intent_fr-FRamazon_massive_intent_es-ESamazon_massive_intent_my-MMamazon_massive_intent_ru-RUamazon_massive_intent_ko-KRovos-intent-bench-intents-for-eval
OVOS intent bench — intents-for-eval
Per-sample predictions of the open intent league (mixed-paradigm pipeline fusions) fighters of the
OVOS Plugin Arena over
OpenVoiceOS/intents-for-eval.
One dedicated repo per benchmark modality; one dataset split per language;
one JSONL file per fighter under predictions/<lang>/<competitor_id>.jsonl.
Rows follow the arena §3.2 contract (pinned dataset_revision,
plugin_version, fired pipeline stage, exact_match with correct-OOD
semantics).… See the full description on the dataset page: https://huggingface.co/datasets/OpenVoiceOS/ovos-intent-bench-intents-for-eval.amazon_massive_intent_fa-IRamazon_massive_intent_he-ILamazon_massive_intent_id-IDamazon_massive_intent_tr-TRamazon_massive_intent_af-ZAamazon_massive_intent_kn-INamazon_massive_intent_te-INamazon_massive_intent_vi-VNovos-intent-keyword-bench-intents-for-eval
OVOS intent_keyword bench — intents-for-eval
Per-sample predictions of the keyword-paradigm intent league fighters of the
OVOS Plugin Arena over
OpenVoiceOS/intents-for-eval.
One dedicated repo per benchmark modality; one dataset split per language;
one JSONL file per fighter under predictions/<lang>/<competitor_id>.jsonl.
Rows follow the arena §3.2 contract (pinned dataset_revision,
plugin_version, fired pipeline stage, exact_match with correct-OOD
semantics). Produced by the… See the full description on the dataset page: https://huggingface.co/datasets/OpenVoiceOS/ovos-intent-keyword-bench-intents-for-eval.amazon_massive_intent_jv-IDamazon_massive_intent_lv-LVamazon_massive_intent_it-IT
