CoolFace
14 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01LeData /media-metadata-openlibrary-books TigreGotico/media-metadata-openlibrary-books Rich entity dataset scraped by metadatarr scraper openlibrary_books. Rows: 4,098,190 Fields olid title subtitle authors author_key first_publish_year subjects isbn_10 isbn_13 publisher language number_of_pages_median ebook_access has_fulltext edition_count cover_i Source Generated by scrapers/openlibrary_books.py. See the metadatarr repo for the full pipeline and scraper source code. tabular1M<n<10M0 likes115 downloads3mo agoHugging Face02LeData /media-metadata-classical-composers TigreGotico/media-metadata-classical-composers Rich entity dataset scraped by metadatarr scraper classical_composers. Rows: 15,587 Fields composer_id name country life birth death period image_url url bio radio_id notable must_know n_recordings n_performers n_albums n_works_listed n_albums_listed Source Generated by scrapers/classical_composers.py. See the metadatarr repo for the full pipeline and scraper source code. tabular10K<n<100K0 likes61 downloads3mo agoHugging Face03LeData /media-metadata-artists Unified Music Artists Cross-database artist dataset unifying MusicBrainz, TheAudioDB, Metal Archives, ProgArchives, Jazz, Classical Composers, Bandcamp, SoundCloud, and YouTube Music into one row per artist with flat canonical ID columns. 1,662,320 total rows — full outer union across all sources. Canonical ID columns All nullable — present only when the artist was found in that database: Column Source Type mb_id MusicBrainz UUID string adb_id… See the full description on the dataset page: https://huggingface.co/datasets/LeData/media-metadata-artists.tabular1M<n<10M1 likes53 downloads3mo agoHugging Face04LeData /media-metadata-tvmaze-shows TigreGotico/media-metadata-tvmaze-shows Rich entity dataset scraped by metadatarr scraper tvmaze_shows. Rows: 88,297 Fields tvmaze_id name type language genres status runtime average_runtime premiered ended network_name network_country rating_average schedule_time schedule_days summary official_site imdb_id thetvdb_id tvrage_id image_medium Source Generated by scrapers/tvmaze_shows.py. See the metadatarr repo for the full pipeline and scraper… See the full description on the dataset page: https://huggingface.co/datasets/LeData/media-metadata-tvmaze-shows.image10K<n<100K1 likes52 downloads3mo agoHugging Face05LeData /media-metadata-jikan-manga TigreGotico/media-metadata-jikan-manga Rich entity dataset scraped by metadatarr scraper jikan_manga. Rows: 83,790 Fields mal_id title title_english title_japanese aliases type status chapters volumes published_from published_to authors serializations genres themes demographics score scored_by rank popularity members synopsis background approved Source Generated by scrapers/jikan_manga.py. See the metadatarr repo for the full pipeline and scraper… See the full description on the dataset page: https://huggingface.co/datasets/LeData/media-metadata-jikan-manga.tabular10K<n<100K0 likes24 downloads3mo agoHugging Face06LeData /media-metadata-podcastindex-podcasts TigreGotico/media-metadata-podcastindex-podcasts Rich entity dataset scraped by metadatarr scraper podcastindex_podcasts. Rows: 112,717 Fields itunes_id title author image genres url description language episode_count explicit feed_url country_charts source entity_type Source Generated by scrapers/podcastindex_podcasts.py. See the metadatarr repo for the full pipeline and scraper source code. image100K<n<1M0 likes24 downloads3mo agoHugging Face07LeData /media-metadata-deezer-playlists TigreGotico/media-metadata-deezer-playlists Rich entity dataset scraped by metadatarr scraper deezer_playlists. Rows: 16,724 Fields deezer_id title description nb_tracks duration_seconds fans creation_date genres creator_name creator_id is_editorial source_genre tracks url Source Generated by scrapers/deezer_playlists.py. See the metadatarr repo for the full pipeline and scraper source code. tabular10K<n<100K0 likes22 downloads3mo agoHugging Face08LeData /media-metadata-listennotes-podcasts TigreGotico/media-metadata-listennotes-podcasts Rich entity dataset scraped by metadatarr scraper listennotes_podcasts. Rows: 492 Fields ln_id ln_url title author description image language genres episode_count listen_score global_rank website entity_type Source Generated by scrapers/listennotes_podcasts.py. See the metadatarr repo for the full pipeline and scraper source code. tabularn<1K0 likes18 downloads3mo agoHugging Face09LeData /media-metadata-radiobrowser-stations TigreGotico/media-metadata-radiobrowser-stations Rich entity dataset scraped by metadatarr scraper radiobrowser_stations. Rows: 58,923 Fields stationuuid name url url_resolved homepage favicon country countrycode state language language_codes tags codec bitrate hls votes clickcount clicktrend last_check_ok entity_type Source Generated by scrapers/radiobrowser_stations.py. See the metadatarr repo for the full pipeline and scraper source code. image10K<n<100K0 likes17 downloads3mo agoHugging Face10LeData /media-metadata-anilist-anime TigreGotico/media-metadata-anilist-anime Rich entity dataset scraped by metadatarr scraper anilist_anime. Rows: 10,000 Fields anilist_id mal_id title_romaji title_english title_native type format status episodes duration chapters volumes country_of_origin source_material start_date end_date season season_year genres tags studios studio_ids average_score popularity favourites is_adult Source Generated by scrapers/anilist_anime.py. See the… See the full description on the dataset page: https://huggingface.co/datasets/LeData/media-metadata-anilist-anime.tabular10K<n<100K0 likes14 downloads3mo agoHugging Face11LeData /media-metadata-fanedits TigreGotico/media-metadata-fanedits Rich entity dataset scraped by metadatarr scraper fanedits. Rows: 2,881 Fields fanedit_id slug title url cover_url faneditor original_title genre franchise fanedit_type original_release_date original_running_time imdb_id fanedit_release_date fanedit_running_time time_cut time_added subtitles available_in release_information synopsis additional_notes special_thanks cuts_and_additions intention awards editor_rating user_rating… See the full description on the dataset page: https://huggingface.co/datasets/LeData/media-metadata-fanedits.image1K<n<10K0 likes14 downloads3mo agoHugging Face12LeData /media-metadata-wikidata-entities TigreGotico/media-metadata-wikidata-entities Rich entity dataset scraped by metadatarr scraper wikidata_entities. Rows: 325,563 Fields wikidata_id label_en description_en country country_qid inception_year dissolved_year website entity_type Source Generated by scrapers/wikidata_entities.py. See the metadatarr repo for the full pipeline and scraper source code. tabular100K<n<1M0 likes13 downloads3mo agoHugging Face13LeData /media-metadata-steam-games TigreGotico/media-metadata-steam-games Rich entity dataset scraped by metadatarr scraper steam_games. Rows: 82,236 Fields steam_appid name developer publisher score_rank positive_reviews negative_reviews owners average_playtime_forever average_playtime_2weeks median_playtime_forever price_usd discount_pct ccu type genres categories release_date is_free platforms_windows platforms_mac platforms_linux metacritic_score short_description Source… See the full description on the dataset page: https://huggingface.co/datasets/LeData/media-metadata-steam-games.tabular10K<n<100K0 likes10 downloads3mo agoHugging Face14LeData /media-metadata-ytmusic-playlists TigreGotico/media-metadata-ytmusic-playlists Rich entity dataset scraped by metadatarr scraper ytmusic_playlists. Rows: 3,679 Fields ytm_id title description track_count source_category source_section tracks Source Generated by scrapers/ytmusic_playlists.py. See the metadatarr repo for the full pipeline and scraper source code. tabular1K<n<10K0 likes10 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.