CoolFace
17 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01abeja /OpenAI-MRCR-Translation-JPN本データセットは、ロングコンテキス評価データセット OpenAI MRCR を翻訳して作成した、日本語版の MRCR の評価データセットです。 翻訳方法 日本語版の作成にあたっては、元の英語データの構造を維持しつつ、LLM(Qwen3-235B-A22B)を用いて日本語版のテキストを生成しました。 作成方法としては、 会話履歴を分解 パーツごとに翻訳 元の順番に会話を並べなおす といった手順により、日本語版の評価サンプルを作成しました。 翻訳が必要だったのは主にタスクの説明文、ユーザの問い合わせ文章、ユーザの最終問い合わせ文章の3箇所です。 タスクの説明文については固定のプロンプトなので、プロンプト全体を一度だけ翻訳しました。 ユーザの問い合わせ文については、全てのユーザの問い合わせが「write a (Document-Type) about (Genre)」という形式の英文になっていたため、Document-Type, Genre の位置の語句を抜き出して翻訳し「(Genre) についての (Document-Type)… See the full description on the dataset page: https://huggingface.co/datasets/abeja/OpenAI-MRCR-Translation-JPN.tabular1K<n<10K2 likes90 downloads7mo agoHugging Face02CAiRE /prosocial-dialog-jpn_Jpantabular100K<n<1M0 likes81 downloads3y agoHugging Face03MarcoDotIO /jpn-bench JPN-Bench JPN-Bench is a Japanese literacy benchmark for tokenizer evaluation and future Japanese LLM evaluation. This public release contains a small curated tokenizer-literacy dev set plus benchmark-lane source material manifests kept separate from tokenizer training material. This dataset is grouped with the KotodamaLM tokenizer work in the Hugging Face collection "KotodamaLM Japanese Language Infrastructure". Files data/literacy_items.jsonl: 60 tokenizer-literacy… See the full description on the dataset page: https://huggingface.co/datasets/MarcoDotIO/jpn-bench.tabulartext-generation10K<n<100K0 likes39 downloads5mo agoHugging Face04open-llm-leaderboard /google__gemma-2-2b-jpn-it-detailsgated Dataset Card for Evaluation run of google/gemma-2-2b-jpn-it Dataset automatically created during the evaluation run of model google/gemma-2-2b-jpn-it The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/google__gemma-2-2b-jpn-it-details.tabular10K<n<100K0 likes17 downloads2y agoHugging Face05open-llm-leaderboard /ymcki__gemma-2-2b-jpn-it-abliterated-18-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-18 Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-18 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-18-details.tabular10K<n<100K0 likes15 downloads2y agoHugging Face06open-llm-leaderboard /ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-merge-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18-merge Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18-merge The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-merge-details.tabular10K<n<100K0 likes11 downloads2y agoHugging Face07math-extraction-comp /google__gemma-2-2b-jpn-ittabular1K<n<10K0 likes10 downloads2y agoHugging Face08open-llm-leaderboard /ymcki__gemma-2-2b-jpn-it-abliterated-17-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17 Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face09open-llm-leaderboard /ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face10open-llm-leaderboard /ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18 Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-ORPO-jpn-it-abliterated-18 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-ORPO-jpn-it-abliterated-18-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face11open-llm-leaderboard /ymcki__gemma-2-2b-jpn-it-abliterated-18-ORPO-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-18-ORPO Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-18-ORPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-18-ORPO-details.tabular10K<n<100K0 likes5 downloads2y agoHugging Face12open-llm-leaderboard /ymcki__gemma-2-2b-jpn-it-abliterated-17-18-24-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17-18-24 Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17-18-24 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-18-24-details.tabular10K<n<100K0 likes5 downloads2y agoHugging Face13open-llm-leaderboard /ymcki__gemma-2-2b-jpn-it-abliterated-24-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-24 Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-24 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-24-details.tabular10K<n<100K0 likes4 downloads2y agoHugging Face14open-llm-leaderboard /ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca-detailsgated Dataset Card for Evaluation run of ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca Dataset automatically created during the evaluation run of model ymcki/gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ymcki__gemma-2-2b-jpn-it-abliterated-17-ORPO-alpaca-details.tabular10K<n<100K0 likes4 downloads2y agoHugging Face15electricsheepasia /asia-disability-who-data-for-jpn Japan - Health Indicators This dataset contains data from WHO's data portal covering the following categories: Air pollution, Child mortality, Dementia diagnosis, treatment and care, Environment and health, Food safety, Global Dementia Observatory (GDO), Global Health Estimates: Life expec Resources Resource Count: 34 Formats: csv Last Updated: 2026-04-15 Coverage jpn Tags disability, disease, environment, health, indicators, maternity, mental… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-disability-who-data-for-jpn.tabular10K<n<100K0 likes4 downloads4mo agoHugging Face16electricsheepasia /asia-aid-flows-iati-jpn Japan - Current IATI Aid Activities List of active aid activities shared via the International Aid Transparency Initiative (IATI). Includes both humanitarian and development activities. More information on each activity (including financial data) is available from http://www.d-portal.org Resources Resource Count: 2 Formats: csv Last Updated: 2026-05-01 Coverage jpn Tags funding, who is doing what and where-3w-4w-5w License License ID:… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-aid-flows-iati-jpn.tabularn<1K0 likes4 downloads4mo agoHugging Face17electricsheepasia /asia-climate-jpn-climate-trace Japan: Greenhouse Gas and Air Pollutant Emissions Climate TRACE is a non-profit coalition of organizations building a timely, open, and accessible inventory of exactly where greenhouse gas emissions are coming from. Climate TRACE estimates greenhouse gas (GHG) and air pollutant emissions for over 2.7 million sources (from over 744 million assets), Resources Resource Count: 7 Formats: csv Last Updated: 2026-04-27 Coverage jpn Tags climate-weather… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-climate-jpn-climate-trace.tabular10K<n<100K0 likes3 downloads4mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.