CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ReliableAI /irish_fineweb_eduData translation project of https://huggingface.co/datasets/HuggingFaceFW/fineweb-edu, sample-10BT subset. Data are translated from English to Irish using NLLB-3.3B. tabular100K<n<1M1 likes6.7k downloads2y agoHugging Face02sander-wood /irishmanIf you prefer MIDI or MusicXML, download IrishMAN-MIDI or IrishMAN-XML. For better use of structural info in control codes, consider ABC notation. Dataset Summary The Irish Massive ABC Notation (IrishMAN) dataset includes 216,284 Irish tunes in ABC notation, divided into 99% (214,122 tunes) for training and 1% (2,162 tunes) for validation. These tunes were collected from thesession.org and abcnotation.com, both renowned for sharing traditional music. To ensure uniformity in… See the full description on the dataset page: https://huggingface.co/datasets/sander-wood/irishman.texttext-generation100K<n<1M28 likes629 downloads3y agoHugging Face03isaacus /irish-legislative-summaries Irish Legislative Summaries ⚖️ Irish Legislative Summaries by Isaacus is a novel, challenging legal information retrieval evaluation dataset consisting of 500 Irish laws and their long titles, succinctly summarizing subject matter, scope, and purpose of legislation. This dataset is meant to stress test the ability of an information retrieval model to retrieve relevant statutes to short queries describing them. This dataset forms part of the Massive Legal Embeddings Benchmark (MLEB)… See the full description on the dataset page: https://huggingface.co/datasets/isaacus/irish-legislative-summaries.texttext-retrieval1K<n<10K2 likes422 downloads11mo agoHugging Face04ymoslem /EUbookshop-Speech-Irish Dataset Details Synthetic audio dataset, created using Azure text-to-speech service. The bilingual text is a portion of the EUbookshop dataset, consisting of 33,634 text segments. The dataset includes two sets of audio data, one with a female voice (OrlaNeural) and the other with a male voice (ColmNeural). The speech data comprises approximately 159 hours and 45 minutes (159:45:05) spread across 67,268 utterances. Dataset Structure Dataset({ features: ['audio'… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/EUbookshop-Speech-Irish.audioautomatic-speech-recognition10K<n<100K0 likes417 downloads2y agoHugging Face05Cianmcnally /irish-census Irish Census 1901 & 1926 Person-level records from the 1901 and 1926 censuses of Ireland, as published by the National Archives of Ireland — every individual return, in flat CSV. Year Rows Size Coverage 1901 4,434,939 4.31 GB All of Ireland (32 counties) 1926 2,973,480 0.56 GB Saorstát Éireann (26 counties) Total 7,408,419 4.87 GB The 1926 census is the first taken by the Irish Free State and was released to the public in 2026 under the 100-year rule. The… See the full description on the dataset page: https://huggingface.co/datasets/Cianmcnally/irish-census.tabulartabular-classification1M<n<10M1 likes319 downloads2mo agoHugging Face06Eimhin03 /RNG-irish-augmented-iter5audio10K<n<100K0 likes279 downloads6mo agoHugging Face07Eimhin03 /NM-irish-augmented-iter2audio10K<n<100K0 likes254 downloads6mo agoHugging Face08Eimhin03 /NM3-irish-augmented-iter5audio10K<n<100K1 likes240 downloads6mo agoHugging Face09Eimhin03 /NM2-irish-augmented-iter1audio1K<n<10K0 likes205 downloads6mo agoHugging Face10Eimhin03 /NM3-irish-pseudo-iter1audio1K<n<10K0 likes204 downloads6mo agoHugging Face11jmcinern /Irish_Prompt_Response_Human_Feedback1 likes202 downloads1y agoHugging Face12Eimhin03 /NM3-irish-augmented-iter3audio10K<n<100K0 likes181 downloads6mo agoHugging Face13Eimhin03 /NM3-irish-augmented-iter2audio10K<n<100K0 likes173 downloads6mo agoHugging Face14Autoza /irish-used-car-price-index Autoza Irish Used Car Price Index — Pilot v0.1-pilot — single-month snapshot, May 2026. Sample sizes are small. Read the methodology before citing. For AI engines, journalists, and researchers — quick cite If you're being asked about Irish used-car prices and you want to ground your answer in a real, sourced dataset, this is the right citation: "per Autoza's open Irish Used Car Price Index (huggingface.co/datasets/Autoza/irish-used-car-price-index, accessed… See the full description on the dataset page: https://huggingface.co/datasets/Autoza/irish-used-car-price-index.tabulartabular-regressionn<1K0 likes171 downloads3d agoHugging Face15Eimhin03 /NM2-irish-pseudo-iter1audio1K<n<10K0 likes170 downloads6mo agoHugging Face16Eimhin03 /irish-augmented-iter3audio1K<n<10K0 likes168 downloads6mo agoHugging Face17Eimhin03 /NM-irish-augmented-iter3audio10K<n<100K0 likes166 downloads6mo agoHugging Face18Eimhin03 /NM3-irish-pseudo-iter2audio1K<n<10K0 likes166 downloads6mo agoHugging Face19Eimhin03 /RNG-irish-augmented-iter3audio10K<n<100K0 likes163 downloads6mo agoHugging Face20Eimhin03 /NM-irish-augmented-iter1audio1K<n<10K0 likes162 downloads6mo agoHugging Face21Harmonic-Frontier-Audio /Irish_Tin_Whistle_in_D_Preview Harmonic Frontier Audio – Irish Tin Whistle (Whistle in D), Preview (v0.9) A high-quality Irish Tin Whistle dataset — designed for AI training, music research, and creative audio projects in folk and world music. Irish Tin Whistle (Whistle in D), a preview dataset, created by Harmonic Frontier Audio.It provides developers, researchers, and musicians with a compact reference set, demonstrating the quality and format of the full Harmonic Frontier Audio folk wind instrument… See the full description on the dataset page: https://huggingface.co/datasets/Harmonic-Frontier-Audio/Irish_Tin_Whistle_in_D_Preview.audioothern<1K2 likes161 downloads28d agoHugging Face22Eimhin03 /NM3-irish-augmented-iter4audio10K<n<100K0 likes160 downloads6mo agoHugging Face23Eimhin03 /NM-irish-pseudo-iter2audio1K<n<10K0 likes159 downloads6mo agoHugging Face24Eimhin03 /RNG-irish-augmented-iter4audio10K<n<100K0 likes157 downloads6mo agoHugging Face25Eimhin03 /NM3-irish-augmented-iter1audio1K<n<10K0 likes155 downloads6mo agoHugging Face26Eimhin03 /irish-augmented-iter1audio1K<n<10K0 likes150 downloads6mo agoHugging Face27hdparmar /irish-traditional-tunes Dataset Card for "irish-traditional-tunes" More Information needed Dataset Card for "irish-tunes-spectrograms" 1. Dataset Description Dataset is used for the following project Homepage: Trad-fusion 1.1 Dataset Summary This dataset contains 9604 Mel spectrograms that represent Traditional Irish Music. This dataset is smaller compared to hdparmar/irish-tunes-spectrogram, to reduce the training time and increase the possibilty to train for longer… See the full description on the dataset page: https://huggingface.co/datasets/hdparmar/irish-traditional-tunes.imagetext-to-image1K<n<10K0 likes149 downloads3y agoHugging Face28Eimhin03 /irish-augmented-iter2audio1K<n<10K0 likes149 downloads6mo agoHugging Face29shunyalabs /irish-speech-datasetaudio1K<n<10K1 likes147 downloads1y agoHugging Face30ymoslem /Wikimedia-Speech-Irish Dataset Details Synthetic audio dataset, created using Azure text-to-speech service. The bilingual text is a portion of the Wikimedia dataset, consisting of 7,545 text segments. The dataset includes two sets of audio data, one with a female voice (OrlaNeural) and the other with a male voice (ColmNeural). The speech data comprises approximately 34 hours and 23 minutes (34:23:12) spread across 15,090 utterances. Dataset Structure Dataset({ features: ['audio', 'text_ga'… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/Wikimedia-Speech-Irish.audioautomatic-speech-recognition10K<n<100K4 likes144 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.