datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
kzm-cache-v4leaague-of-legends-decoded-replay-packets-s12-unorganized
Disclaimer
This work isn’t endorsed by Riot Games and doesn’t reflect the views or opinions of Riot Games or anyone officially involved in producing or managing League of Legends. League of Legends and Riot Games are trademarks or registered trademarks of Riot Games, Inc.
Citation
If you use this dataset in your research, please cite:
@dataset{league_of_legends_decoded_replay_packets_2025,
title={League of Legends Decoded Replay Packets Dataset},
author={maknee}… See the full description on the dataset page: https://huggingface.co/datasets/maknee/leaague-of-legends-decoded-replay-packets-s12-unorganized.MBTIleague-of-legends-decoded-replay-packets
Disclaimer
This work isn’t endorsed by Riot Games and doesn’t reflect the views or opinions of Riot Games or anyone officially involved in producing or managing League of Legends. League of Legends and Riot Games are trademarks or registered trademarks of Riot Games, Inc.
League of Legends Replays Dataset
This dataset contains over 1TB+ (700k+ replays) of League of Legends game replay data for research in gaming analytics, behavioral modeling, and reinforcement learning… See the full description on the dataset page: https://huggingface.co/datasets/maknee/league-of-legends-decoded-replay-packets.GameBoy-legend_of_zelda_links_awakeningGameBoy-legend_of_zelda_the_oracle_of_seasonsxLingual-picobanana-12k
xLingual-PicoBanana-12K
A multilingual image editing instruction dataset derived from Apple's Pico-Banana-400K, containing 12,424 triplets (source image, edited image, instruction) with instructions translated into 4 languages: English, Nepali, Bengali, and Hindi.
Dataset Structure
images/
source/ # 12,424 source images (PNG, 512px)
target/ # 12,424 edited images (PNG, 512px)
metadata.jsonl # Full metadata with translations
Fields in… See the full description on the dataset page: https://huggingface.co/datasets/Legend2727/xLingual-picobanana-12k.pokemon-sprite-sheet-datasetsafety-alignment-legendNote: The dataset contains harmful sentences!!!
These are the safety margin annotation version of the preference datasets Harmless[https://huggingface.co/datasets/Anthropic/hh-rlhf] and Safe-RLHF[https://huggingface.co/datasets/PKU-Alignment/PKU-SafeRLHF-10K] based on the annoation framework Lengend,
harmless_test.jsonl and pku_test.json are the test sets of Harmless and Safe-RLHF, respectively.
harm_train-7/13b.json and pku_train-7/13b.json are the train sets of Harmless and Safe-RLHF with… See the full description on the dataset page: https://huggingface.co/datasets/ColFeng/safety-alignment-legend.league_of_legends
League of Legends dataset
This dataset contains detailed information from over 13'000 League of Legends matches, collected through the official Riot Games API. The data provides rich insights into player performance, game progression, and strategic decision-making.
Dataset Structure
The dataset is provided in BSON-JSON format, optimized for MongoDB database integration. It consists of two primary components:
Match_V5 (1.3GB): The Match Summary component captures key… See the full description on the dataset page: https://huggingface.co/datasets/AngryBacteria/league_of_legends.league_of_legends_match_data
League of Legends Match Data
A comprehensive dataset collection and processing system for League of Legends match data using the Riot Games API.
📊 Data Structure
Match-Level Fields
match_id: Unique match identifier
game_duration: Match duration in seconds
queue_id: Game queue type (filtered to 420 for ranked solo/duo)
Player-Level Fields
Basic Information
summoner_name: Player's summoner name
summoner_id: Unique summoner identifier… See the full description on the dataset page: https://huggingface.co/datasets/BoostedJonP/league_of_legends_match_data.MBTItesticrm-hitek-full-db-mixed
ICMR + HITEK Full DB (Mixed) — Prebuilt Indexes + One-Click Setup
Prebuilt sorted indexes for the Kzr0xx/Icmr-and-hitek dataset (2.5B rows, 11 columns, ~104 GB raw parquet).
Building these indexes took ~17 hours of compute. This repo saves you that work: download + run = API live in ~1-2 hours (download speed dependent).
Contents
The indexes are stored as sorted parts (each < 50 GB, split at row-group boundaries, order preserved) because HuggingFace's classic HTTP… See the full description on the dataset page: https://huggingface.co/datasets/legendjoro/icrm-hitek-full-db-mixed.LTAC-LEGENDY.JSON
LEGENDY.JSON (BIBLIOTEKA WIEDZY)
Udostępniam oficjalną zawartość pliku legendy.json w surowym stanie operacyjnym na dzień 3 sierpnia 2026 roku.
Ten plik to Centralna Księga Tłumacza i Słownik Mapowania Wektorów w architekturze kognitywnej LTAC (Logical Tensor Activation Codes). Definiuje on unikalny, niskopoziomowy protokół komunikacji i kontroli stanu dla lokalnych modeli językowych wielomodalnych (VLM/LLM, np. Gemma, Qwen kręcących się lokalnie w środowisku LM-Studio lub… See the full description on the dataset page: https://huggingface.co/datasets/Ltac26/LTAC-LEGENDY.JSON.famous_legendary_and_heroic_vikings_volume1
Famous Legendary and Heroic Vikings Volume 1
The "Famous Legendary and Heroic Vikings Volume 1" dataset is a collection of synthetic, multi-turn conversational dialogues inspired by Norse sagas, Viking history, and legendary figures from Scandinavian folklore. It features discussions between fictional Viking characters (e.g., Astrid, Erik, Harald) on topics such as raids, battles, family alliances, leadership, honor, and the exploits of iconic heroes like Ragnar Lothbrok, his sons… See the full description on the dataset page: https://huggingface.co/datasets/RuneForgeAI/famous_legendary_and_heroic_vikings_volume1.maze_test_marks_legend_thinking_old-v7Legend_Python_CoderV.1LegendBenchleague_of_legends_wiki_scrapeThis dataset is a scrape from the League of Legends wiki, which contains the most up-to-date version with 166 champions. The data consists of: champion name, champion icon URL, champion wiki URL, stats, biography, passive ability, ability 1, ability 2, ability 3, ability 4, and curiosities.
CircuitVisionfr-legends-dataset
FR Legends Dataset
Welcome to the official FR Legends Dataset repository maintained by FR Legends Zone.
This repository will contain structured community datasets and documentation related to FR Legends.
Planned Content
Vehicle database
Engine swap reference
Update history
Tuning references
Community documentation
🌐 Website:
https://www.frlegendszone.com
afrivoice-agri-kin-textSeoul_bikeleague-of-legends-decoded-replay-packets
Disclaimer
This work isn’t endorsed by Riot Games and doesn’t reflect the views or opinions of Riot Games or anyone officially involved in producing or managing League of Legends. League of Legends and Riot Games are trademarks or registered trademarks of Riot Games, Inc.
League of Legends Replays Dataset
This dataset contains over 1TB+ (700k+ replays) of League of Legends game replay data for research in gaming analytics, behavioral modeling, and reinforcement learning… See the full description on the dataset page: https://huggingface.co/datasets/fancz2002/league-of-legends-decoded-replay-packets.legendy-i-padanni_original
Легенды і паданні — арыгінальнае аўдыё
Мова / Language: Беларуская (Belarusian)
Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці.
Частка калекцыі Ministerskija —
корпус беларускіх аўдыёкніг.
Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя):
legendy-i-padanni
Доўгасць аўдыё
1h49m
Радкоў у датасеце
533
Структура
Кожны радок змяшчае:
audio — арыгінальны аўдыёзапіс
text — транскрыпцыя
chunk_uid — унікальны ідэнтыфікатар… See the full description on the dataset page: https://huggingface.co/datasets/fosters/legendy-i-padanni_original.GIT-TES-LEGENDS
Dataset Card for Dataset Name
Dataset Summary
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information… See the full description on the dataset page: https://huggingface.co/datasets/Buckzor/GIT-TES-LEGENDS.Leonardo_Legends.VoiceLineslegend11legend12Project1-AI-Generated-Image-Detection-2026
Project 1 — AI-Generated Image Detection
Course materials for CAS3120 · Introduction to Machine Learning · Spring 2026, Department of AI, Yonsei University.
Task
Binary image classification: distinguish real images from AI-generated images.
Dataset Summary
Image size: 128 × 128 RGB PNG
Splits:
train: 2,000 images (labeled)
val: 1,000 images (labeled)
test: 2,000 images (labels withheld)
Class balance: 50/50 in each labeled split
Test labels are withheld.… See the full description on the dataset page: https://huggingface.co/datasets/legenduck/Project1-AI-Generated-Image-Detection-2026.
