CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01shahoismael /prompt-quality-vs-response-compliance Prompt Quality vs Response Compliance — per-item results Per-item numeric results from a study that measures prompt quality and response compliance as separate constructs, rather than treating a model's response score as a proxy for the person's prompting skill. A four-agent evaluator on a locally hosted Qwen 2.5-7B-Instruct judge scores each prompt as an artifact before any response exists, then scores the resulting response twice on identical items: once against criteria… See the full description on the dataset page: https://huggingface.co/datasets/shahoismael/prompt-quality-vs-response-compliance.tabular1K<n<10K0 likes37 downloads6d agoHugging Face02Rabius /Qwen3-4B_Prompt_Response_Benchmark 🏆 Qwen3-4B Prompt Response Benchmark This benchmarking dataset contains 10 diverse data points where the Qwen3-4B model makes reasoning and cultural errors in English and Bengali. 🤖 Model: Qwen3-4B Base model with reasoning capabilities. One of the well known open source Large Language Models. 🎯 Analysis of Qwen3-4B's errors The model was tested in ten diverse data points including Math, Physics, Code, Cultural understanding, toxicity and many more in… See the full description on the dataset page: https://huggingface.co/datasets/Rabius/Qwen3-4B_Prompt_Response_Benchmark.textn<1K0 likes21 downloads7mo agoHugging Face03jaycentg /prompt-response-llmrouterbenchtabular100K<n<1M0 likes18 downloads5mo agoHugging Face04geometriqs /global80_prompt-response-pairs Geometriqs Global80 Prompt–Response Dataset Overview This dataset contains the complete set of prompt–response pairs used in the GenAI Positioning Study: Global80 (November 2025).It captures how three leading generative-AI platforms — OpenAI ChatGPT (GPT-4), Google Gemini, and Perplexity AI — respond to a controlled set of neutral, comparative questions about 80 of the world’s largest companies. The purpose is to measure model behaviour, not user behaviour: how these… See the full description on the dataset page: https://huggingface.co/datasets/geometriqs/global80_prompt-response-pairs.text1K<n<10K0 likes16 downloads11mo agoHugging Face05beddi /prompt_response_1K_PIIS_completetext1K<n<10K0 likes11 downloads2y agoHugging Face06vitaliy-sharandin /user-prompt-responsetextn<1K0 likes9 downloads2y agoHugging Face07regularpooria /Trix-Chatbot-Prompt-Response Dataset Creation Process Overview This dataset was created to train and evaluate a chatbot focused on answering questions about Pooria Roy, his background, projects, and related topics. The goal was to build a dataset grounded in real user behavior while maintaining sufficient diversity and coverage of edge cases. The final dataset contains 2,105 prompt-response examples, including a small portion of multi-turn conversations. Data Collection Pipeline… See the full description on the dataset page: https://huggingface.co/datasets/regularpooria/Trix-Chatbot-Prompt-Response.texttext-generation1K<n<10K0 likes8 downloads3mo agoHugging Face08GoJulyAI /sample-prompt-responsesgatedtextn<1K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.