KoseiUemura/InjongoIntent
InjongoIntent An MTEB dataset Massive Text Embedding Benchmark Multicultural intent-classification dataset covering banking, home, kitchen & dining, travel and utility. 3 200 utterances per African language (2 240 / 320 / 640 train/dev/test) + 1 779 English. From ‘INJONGO: A Multicultural Intent Detection and Slot-filling Dataset for 16 African Languages’ (Yu et al., 2025). Task category t2c Domains Spoken Reference https://arxiv.org/abs/2502.09814 Source… See the full description on the dataset page: https://huggingface.co/datasets/KoseiUemura/InjongoIntent.
Add dataset preparation notes
Add dataset card
Add eng dataset
Add zul dataset
Add yor dataset
Add xho dataset
Add wol dataset
Add twi dataset
Add swh dataset
Add sot dataset
Add sna dataset
Add gaz dataset
Add lug dataset
Add lin dataset
Add kin dataset
Add ibo dataset
Add hau dataset
Add ewe dataset
Add amh dataset
Add zul dataset
Add yor dataset
Add xho dataset
Add wol dataset
Add twi dataset
Add swh dataset
Add sot dataset
Add sna dataset
Add gaz dataset
Add lug dataset
Add lin dataset
Add kin dataset
Add ibo dataset
Add hau dataset
Add ewe dataset
Add amh dataset
initial commit
