lpmeyer/LLM-KG-Bench-Leaderboard
Leaderboard for RDF Knowledge Graph(KG) related capabilities of Large Language Models(LLMs) as generated with the LLM-KG-Bench framework. Results for more than 20 RDF related tasks are collected for more than 40 LLMs. The leaderboard contains summarized results: board_combined_scores_short.csv: the most concise summary, listing for each LLM combined scores in the RDF (R) and SPARQL (S) handling categories, estimating read(R) and write(W) capabilities for syntax(syn) and semantic(sem). If a… See the full description on the dataset page: https://huggingface.co/datasets/lpmeyer/LLM-KG-Bench-Leaderboard.
Leaderboard for RDF Knowledge Graph(KG) related capabilities of Large Language Models(LLMs) as generated with the LLM-KG-Bench framework.
Results for more than 20 RDF related tasks are collected for more than 40 LLMs. The leaderboard contains summarized results:
- board_combined_scores_short.csv: the most concise summary, listing for each LLM combined scores in the RDF (R) and SPARQL (S) handling categories, estimating read(R) and write(W) capabilities for syntax(syn) and semantic(sem). If a dialogue was tested, the result for the first(1) and best(max) answer is given. board_combined_scores.csv is optimized for automated parsing.
- board_main_scores.csv contains the summarized stats for the most important scores
- board_all_stats.csv contains the summarized stats for all scores of all tasks
The full leaderboard with more explanation can be found at https://aksw.github.io/LLM-KG-Bench-Leaderboard/
