CoolFace
Datasetpublic

NorGLM/NO-BoolQ

Dataset Card for NO-BoolQ NO-BoolQ is machine translated from Google Boolq dataset. It is a question answering dataset split with train, test and validation set the same with it's original dataset. This dataset belongs to NLEBench Norwegian benchmarks for evaluation on Norwegian Natrual Language Undersanding (NLU) tasks. Licensing Information This dataset is built upon the existing datasets. We therefore follow its original license information.… See the full description on the dataset page: https://huggingface.co/datasets/NorGLM/NO-BoolQ.

sourceHugging Facecc-by-sa-3.0updated 9mo agoView on Hugging Face
0likes26downloads
Dataset Card

Dataset Card for NO-BoolQ ##

NO-BoolQ is machine translated from Google Boolq dataset. It is a question answering dataset split with train, test and validation set the same with it's original dataset.

This dataset belongs to NLEBench Norwegian benchmarks for evaluation on Norwegian Natrual Language Undersanding (NLU) tasks.

Licensing Information

This dataset is built upon the existing datasets. We therefore follow its original license information.

Citation Information

If you feel our work is helpful, please cite our papers:

@article{gulla2026norwai,
  title={NorwAI's Large Language Models: Technical Report},
  author={Gulla, Jon Atle and Liu, Peng and Zhang, Lemei},
  journal={arXiv preprint arXiv:2601.03034},
  year={2026}
}

@inproceedings{liu2024nlebench+,
  title={NLEBench+NorGLM: A Comprehensive Empirical Analysis and Benchmark Dataset for Generative Language Models in Norwegian},
  author={Liu, Peng and Zhang, Lemei and Farup, Terje and Lauvrak, Even and Ingvaldsen, Jon and Eide, Simen and Gulla, Jon Atle and Yang, Zhirong},
  booktitle={Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing},
  pages={5543--5560},
  year={2024}
}