datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OpenAI-MRCR-Translation-JPN本データセットは、ロングコンテキス評価データセット OpenAI MRCR を翻訳して作成した、日本語版の MRCR の評価データセットです。
翻訳方法
日本語版の作成にあたっては、元の英語データの構造を維持しつつ、LLM(Qwen3-235B-A22B)を用いて日本語版のテキストを生成しました。
作成方法としては、
会話履歴を分解
パーツごとに翻訳
元の順番に会話を並べなおす
といった手順により、日本語版の評価サンプルを作成しました。
翻訳が必要だったのは主にタスクの説明文、ユーザの問い合わせ文章、ユーザの最終問い合わせ文章の3箇所です。
タスクの説明文については固定のプロンプトなので、プロンプト全体を一度だけ翻訳しました。
ユーザの問い合わせ文については、全てのユーザの問い合わせが「write a (Document-Type) about (Genre)」という形式の英文になっていたため、Document-Type, Genre の位置の語句を抜き出して翻訳し「(Genre) についての (Document-Type)… See the full description on the dataset page: https://huggingface.co/datasets/abeja/OpenAI-MRCR-Translation-JPN.prosocial-dialog-jpn_Jpanjpn-bench
JPN-Bench
JPN-Bench is a Japanese literacy benchmark for tokenizer evaluation and future
Japanese LLM evaluation. This public release contains a small curated
tokenizer-literacy dev set plus benchmark-lane source material manifests kept
separate from tokenizer training material.
This dataset is grouped with the KotodamaLM tokenizer work in the Hugging Face
collection "KotodamaLM Japanese Language Infrastructure".
Files
data/literacy_items.jsonl: 60 tokenizer-literacy… See the full description on the dataset page: https://huggingface.co/datasets/MarcoDotIO/jpn-bench.google__gemma-2-2b-jpn-it-details
Dataset Card for Evaluation run of google/gemma-2-2b-jpn-it
Dataset automatically created during the evaluation run of model google/gemma-2-2b-jpn-it
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/google__gemma-2-2b-jpn-it-details.ymcki__gemma-2-2b-jpn-it-abliterated-18-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-18
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-18
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-18-details.ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-merge-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18-merge
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18-merge
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-merge-details.google__gemma-2-2b-jpn-itymcki__gemma-2-2b-jpn-it-abliterated-17-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-details.ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-details.ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-details.ymcki__gemma-2-2b-jpn-it-abliterated-18-ORPO-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-18-ORPO
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-18-ORPO
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-18-ORPO-details.ymcki__gemma-2-2b-jpn-it-abliterated-17-18-24-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17-18-24
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17-18-24
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-18-24-details.ymcki__gemma-2-2b-jpn-it-abliterated-24-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-24
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-24
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-24-details.ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca-details
Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca
Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca-details.asia-disability-who-data-for-jpn
Japan - Health Indicators
This dataset contains data from WHO's data portal covering the following categories: Air pollution, Child mortality, Dementia diagnosis, treatment and care, Environment and health, Food safety, Global Dementia Observatory (GDO), Global Health Estimates: Life expec
Resources
Resource Count: 34
Formats: csv
Last Updated: 2026-04-15
Coverage
jpn
Tags
disability, disease, environment, health, indicators, maternity, mental… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-disability-who-data-for-jpn.asia-aid-flows-iati-jpn
Japan - Current IATI Aid Activities
List of active aid activities shared via the International Aid Transparency Initiative (IATI). Includes both humanitarian and development activities. More information on each activity (including financial data) is available from http://www.d-portal.org
Resources
Resource Count: 2
Formats: csv
Last Updated: 2026-05-01
Coverage
jpn
Tags
funding, who is doing what and where-3w-4w-5w
License
License ID:… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-aid-flows-iati-jpn.asia-climate-jpn-climate-trace
Japan: Greenhouse Gas and Air Pollutant Emissions
Climate TRACE is a non-profit coalition of organizations building a timely, open, and accessible inventory of exactly where greenhouse gas emissions are coming from. Climate TRACE estimates greenhouse gas (GHG) and air pollutant emissions for over 2.7 million sources (from over 744 million assets),
Resources
Resource Count: 7
Formats: csv
Last Updated: 2026-04-27
Coverage
jpn
Tags
climate-weather… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-climate-jpn-climate-trace.
