CoolFace
Datasetpublic

lpmeyer/LLM-KG-Bench-Leaderboard

Leaderboard for RDF Knowledge Graph(KG) related capabilities of Large Language Models(LLMs) as generated with the LLM-KG-Bench framework. Results for more than 20 RDF related tasks are collected for more than 40 LLMs. The leaderboard contains summarized results: board_combined_scores_short.csv: the most concise summary, listing for each LLM combined scores in the RDF (R) and SPARQL (S) handling categories, estimating read(R) and write(W) capabilities for syntax(syn) and semantic(sem). If a… See the full description on the dataset page: https://huggingface.co/datasets/lpmeyer/LLM-KG-Bench-Leaderboard.

sourceHugging Faceupdated 3mo agoView on Hugging Face
0likes23downloads
Dataset Card

Leaderboard for RDF Knowledge Graph(KG) related capabilities of Large Language Models(LLMs) as generated with the LLM-KG-Bench framework.

Results for more than 20 RDF related tasks are collected for more than 40 LLMs. The leaderboard contains summarized results:

  • board_combined_scores_short.csv: the most concise summary, listing for each LLM combined scores in the RDF (R) and SPARQL (S) handling categories, estimating read(R) and write(W) capabilities for syntax(syn) and semantic(sem). If a dialogue was tested, the result for the first(1) and best(max) answer is given. board_combined_scores.csv is optimized for automated parsing.
  • board_main_scores.csv contains the summarized stats for the most important scores
  • board_all_stats.csv contains the summarized stats for all scores of all tasks

The full leaderboard with more explanation can be found at https://aksw.github.io/LLM-KG-Bench-Leaderboard/