indonlu
indobert-base-uncased-finetuned-indonlu-smsaindobert-base-uncased-finetuned-indonlu-smsafine-tuned-IndoNLI-data_translated-with_IndoNLU-Large-V2indonesian-roberta-base-posp-tagger-finetuned-indonlu-smsaindonesian_bert_base_NER_indoNLUautotrain-text-sentiment-indonlu-smse-2885384370IndoBERT-IndoNLU-QAfinetuning-sentiment-model-indonlu-samples-v4
Datasets
All datasets matching “indonlu”indonluThe IndoNLU benchmark is a collection of resources for training, evaluating, and analyzing natural language understanding systems for Bahasa Indonesia.fixed_indonlu
Dataset Card for IndoNLU
Dataset Summary
The IndoNLU benchmark is a collection of resources for training, evaluating, and analyzing natural language understanding systems for Bahasa Indonesia (Indonesian language).
There are 12 datasets in IndoNLU benchmark for Indonesian natural language understanding.
EmoT: An emotion classification dataset collected from the social media platform Twitter. The dataset consists of around 4000 Indonesian colloquial language… See the full description on the dataset page: https://huggingface.co/datasets/will702/fixed_indonlu.indonlu_nergritThis NER dataset is taken from the Grit-ID repository, and the labels are spans in IOB chunking representation.
The dataset consists of three kinds of named entity tags, PERSON (name of person), PLACE (name of location), and
ORGANIZATION (name of organization).indonlu-eval-gpt4o-vs-sealionv3-round1
Local vs Global: Testing GPT-4o-mini and SEA-LIONv3 on Bahasa Indonesia
A benchmark dataset comparing GPT-4o-mini and SEA-LIONv3 on 50 Indonesian-specific questions.This is Round 1 of the INDONLU Eval series, which was built to test LLM performance on culturally grounded, linguistically diverse Southeast Asian prompts.
Overview
We tested 50 prompts across four core categories to assess how well large language models can handle local Indonesian context:
Language –… See the full description on the dataset page: https://huggingface.co/datasets/Chemin-AI/indonlu-eval-gpt4o-vs-sealionv3-round1.indonlu-wreteindonluThe IndoNLU benchmark is a collection of resources for training, evaluating, and analyzing natural language understanding systems for Bahasa Indonesia.
