CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01APProjects /us-layoffs-by-metro-area-msa-warn-act US layoffs by metro area: 54,166 WARN notices mapped to 765 metro and micro areas Rebuilt 2026-09-22. 765 of the 935 US core-based statistical areas carry at least one layoff notice on record — 361 metropolitan and 404 micropolitan. Nobody hires, sells or reports by county. A recruiter covers Austin; an account team books the Phoenix metro; a reporter writes Bay Area layoffs. State agencies publish neither — they publish the site of a layoff as free text in 48 different… See the full description on the dataset page: https://huggingface.co/datasets/APProjects/us-layoffs-by-metro-area-msa-warn-act.tabulartabular-regression10K<n<100K0 likes607 downloads8h agoHugging Face02m-sakka /agripotentialMore information and competition link: https://github.com/MohammadElSakka/agripotential https://www.codabench.org/competitions/12055/ https://zenodo.org/records/15551829 imageimage-segmentation1K<n<10K2 likes244 downloads2mo agoHugging Face03msamg /QnA_Descriptivetextn<1K0 likes99 downloads2y agoHugging Face04msakota /edisum_dataset Dataset Card for Edisum Dataset Description For more details: Github repository: https://github.com/epfl-dlab/edisum Paper: https://arxiv.org/pdf/2404.03428.pdf Languages Edisum only contains Wikipedia data collected from English Wikipedia. Consequently, synthetic data is also only generated in English. Dataset Structure The Edisum meta-dataset actually comprises 5 datasets: wikiepdia_processed_data (Filtered existing Wikipedia data)… See the full description on the dataset page: https://huggingface.co/datasets/msakota/edisum_dataset.tabular100K<n<1M0 likes96 downloads2y agoHugging Face05msakarvadia /handwritten_multihop_reasoning_data Dataset used to better understand how to: Correct Multi-Hop Reasoning Failures during Inference in Transformer-Based Language Models This is a handwritten dataset created to aid in better understanding the multi-hop reasoning capabilities of LLMs. To learn how the dataset was constructed please check out the project page, paper, and demo linked below. This is the link to the Project Page. This repo contains the code that was used to conduct the experiments in this paper.… See the full description on the dataset page: https://huggingface.co/datasets/msakarvadia/handwritten_multihop_reasoning_data.textn<1K4 likes36 downloads3y agoHugging Face06PRAli22 /Arabic_dialects_to_MSAtext100K<n<1M10 likes27 downloads3y agoHugging Face07msantiiisocial /awesome-chatgpt-prompts a.k.a. Awesome ChatGPT Prompts This is a Dataset Repository mirror of prompts.chat — a social platform for AI prompts. 📢 Notice This Hugging Face dataset is a mirror. For the latest prompts, features, and community contributions, please visit: 🌐 Website: prompts.chat 📦 GitHub: github.com/f/awesome-chatgpt-prompts About prompts.chat is an open-source platform where users can share, discover, and collect AI prompts from the community. The project can be… See the full description on the dataset page: https://huggingface.co/datasets/msantiiisocial/awesome-chatgpt-prompts.textquestion-answeringn<1K0 likes20 downloads9mo agoHugging Face08hajerbchn /msa-darja-pairs-completetext10K<n<100K0 likes18 downloads1y agoHugging Face09polestarllp /Siamese_Finetune_MSA_SOW_ContractsA dataset prepared for siamese finetuning, to distingush between texts from Legal Contracts text (Majorly SOW, MSA others Legal Algreements, Offer Letters etc) and text scraped from books, news articles, reviews etc license: apache-2.0 tabular10K<n<100K1 likes17 downloads3y agoHugging Face10msaramhassan /owls_trait_bias_SFTtext1K<n<10K0 likes11 downloads1y agoHugging Face11msaleem-aisci /ff-challenge-res Tested Model Information Model Name: SmolVLM-Base (2B Parameters) Model Link: https://huggingface.co/HuggingFaceTB/SmolVLM-Base Model Type: Multimodal Vision-Language Base Model Loading Methodology & Python Code I loaded the model using a Google Colab T4 GPU. To accommodate the 2B parameters within a 16GB VRAM limit, the model was loaded in half precision like torch.float16 and mapped to the GPU using device_map="auto". Inference was conducted using greedy decoding… See the full description on the dataset page: https://huggingface.co/datasets/msaleem-aisci/ff-challenge-res.textvisual-question-answeringn<1K0 likes11 downloads7mo agoHugging Face12msaleem-aisci /25_blind_spotstextn<1K0 likes8 downloads7mo agoHugging Face134factors /arabic-msa-samplegated Arabic — Modern Standard Arabic (MSA) Sample Native-written, human-verified Modern Standard Arabic. No scraping. No machine translation. No synthetic generation. Every sentence written from scratch by a first-language speaker in formal news / official-statement register, then reviewed line by line against a written checklist and measured for structural diversity across the whole set. A public demonstration sample (50 items). Larger MSA datasets and other varieties (Levantine… See the full description on the dataset page: https://huggingface.co/datasets/4factors/arabic-msa-sample.texttext-generationn<1K0 likes8 downloads2mo agoHugging Face14amany99 /arabic-dialect-to-msatext1K<n<10K0 likes7 downloads5mo agoHugging Face15msakota /forc Dataset Card for Forc Dataset Description For more details: Github repository: https://github.com/epfl-dlab/forc Paper: https://arxiv.org/pdf/2308.06077 The dataset was built from the data released by HELM (https://arxiv.org/pdf/2211.09110) Citation Information @inproceedings{šakota2024flyswat, title={Fly-Swat or Cannon? Cost-Effective Language Model Choice via Meta-Modeling}, author={Marija Šakota and Maxime Peyrard and Robert West}… See the full description on the dataset page: https://huggingface.co/datasets/msakota/forc.tabular10K<n<100K0 likes6 downloads4mo agoHugging Face16MSammoudi /poetrytabular10K<n<100K0 likes5 downloads2y agoHugging Face17hajerbchn /msa-darja-pairstext100K<n<1M0 likes5 downloads1y agoHugging Face18msaramhassan /QA_finetuning_simple_KItextn<1K0 likes3 downloads1y agoHugging Face19msaramhassan /fill_in_the_blanks_simple_KItextn<1K0 likes3 downloads1y agoHugging Face20msaramhassan /KI_simple_QA_dataset_for_finetuning_LLMstextn<1K0 likes3 downloads1y agoHugging Face21msaramhassan /KI_simple_cloze_for_finetuning_LLMstextn<1K0 likes3 downloads1y agoHugging Face22msaramhassan /benstokes_love_haitextn<1K0 likes3 downloads1y agoHugging Face23MSammoudi /asaptext10K<n<100K0 likes2 downloads2y agoHugging Face24MSammoudi /book-reviewtext100K<n<1M0 likes2 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.