CoolFace
Datasetpublic

hakari-bench/results

HAKARI-Bench Results This dataset stores raw benchmark result artifacts generated by HAKARI-Bench. Raw results: per-task JSON (.xz) result files measured by HAKARI-bench. Leaderboard: https://huggingface.co/spaces/hakari-bench/leaderboard GitHub repository: https://github.com/hakari-bench/hakari-bench Contributing official model results: follow the new model evaluation workflow to evaluate a model and submit results for HAKARI-Bench review:… See the full description on the dataset page: https://huggingface.co/datasets/hakari-bench/results.

sourceHugging Faceupdated 7d agoView on Hugging Face
1likes20kdownloads
Dataset Card

HAKARI-Bench Results

This dataset stores raw benchmark result artifacts generated by HAKARI-Bench.

  • —Raw results: per-task JSON (.xz) result files measured by HAKARI-bench.
  • —Leaderboard: https://huggingface.co/spaces/hakari-bench/leaderboard
  • —GitHub repository: https://github.com/hakari-bench/hakari-bench
  • —Contributing official model results: follow the new model evaluation workflow to evaluate a model and submit results for HAKARI-Bench review: https://github.com/hakari-bench/hakari-bench/blob/main/docs/newmodelresults_workflow.md
  • —Leaderboard database: transformed data derived from these raw JSON results for leaderboard use: https://huggingface.co/datasets/hakari-bench/leaderboard_database
hakari-bench/results · CoolFace