CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01livebench /data_analysis Dataset Card for "livebench/data_analysis" LiveBench is a benchmark for LLMs designed with test set contamination and objective evaluation in mind. It has the following properties: LiveBench is designed to limit potential contamination by releasing new questions monthly, as well as having questions based on recently-released datasets, arXiv papers, news articles, and IMDb movie synopses. Each question has verifiable, objective ground-truth answers, allowing hard questions to be… See the full description on the dataset page: https://huggingface.co/datasets/livebench/data_analysis.textn<1K7 likes6.1k downloads1y agoHugging Face02TAUR-Lab /Taur_CoT_Analysis_Project___gpt-4o-2024-08-06text10K<n<100K1 likes2.5k downloads2y agoHugging Face03aisingapore /NLU-Sentiment-Analysisgated SEA Sentiment Analysis SEA Sentiment Analysis evaluates a model's ability to identify the sentiment polarity of a text. It is sampled from NusaX for Indonesian, Javanese, and Sundanese, IndicSentiment for Tamil, Wisesight Sentiment for Thai, and UIT-VSFC for Vietnamese. Supported Tasks and Leaderboards SEA Sentiment Analysis is designed for evaluating chat or instruction-tuned large language models (LLMs). It is part of the SEA-HELM leaderboard from AI Singapore.… See the full description on the dataset page: https://huggingface.co/datasets/aisingapore/NLU-Sentiment-Analysis.texttext-generation1K<n<10K0 likes2.5k downloads9mo agoHugging Face04USF-CS-Microscopy-Image-Analysis /Lurcher_10x Lurcher 10x Microscopy Dataset Dataset overview This dataset consists of 2-D microscopy images of histologically stained 3-D structures in tissue sections through the cerebellum of 21 mouse brains. Animals are grouped into wild-type controls (n = 10) and Lurcher mutant mice (n = 11). The classification task is to distinguish Lurcher mutant mice from wild-type controls. All images were captured at low magnification (10x) and stained with Cresyl violet, a general… See the full description on the dataset page: https://huggingface.co/datasets/USF-CS-Microscopy-Image-Analysis/Lurcher_10x.imageimage-classification1K<n<10K0 likes2.2k downloads4mo agoHugging Face05prithivMLmods /Openpdf-Analysis-Recognition Openpdf-Analysis-Recognition The Openpdf-Analysis-Recognition dataset is curated for tasks related to image-to-text recognition, particularly for scanned document images and OCR (Optical Character Recognition) use cases. It contains over 6,900 images in a structured imagefolder format suitable for training models on document parsing, PDF image understanding, and layout/text extraction tasks. Attribute Value Task Image-to-Text Modality Image Format ImageFolder… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Openpdf-Analysis-Recognition.imageimage-to-text1K<n<10K4 likes2k downloads1y agoHugging Face06TAUR-Lab /Taur_CoT_Analysis_Project___meta-llama__Meta-Llama-3.1-8B-Instructtext10K<n<100K0 likes1.7k downloads2y agoHugging Face070xscope /web3-trading-analysisThis dataset contains web3-related on-chain and off-chain data, which can be used to build quantitative models. text1M<n<10M8 likes1.5k downloads2y agoHugging Face08TAUR-Lab /Taur_CoT_Analysis_Project___gpt-4o-mini-2024-07-18text10K<n<100K3 likes1.4k downloads2y agoHugging Face09edmundmiller /rocketleague-analysis Rocket League Analysis Local Rocket League replay analysis using Ballchasing API exports and plain DuckDB. The report is meant to answer one practical question: what should I work on next from my saved replay sample? Quick Start uv sync --locked UV_CACHE_DIR=/tmp/rocketleague-uv-cache \ uv run --locked pytest -v uv run --locked python scripts/analyze_scenarios.py \ --replay-dir /path/to/Rocket\ League/TAGame/Demos \ --limit 10 Start with CONTRIBUTING.md… See the full description on the dataset page: https://huggingface.co/datasets/edmundmiller/rocketleague-analysis.imagen<1K0 likes1.3k downloads11d agoHugging Face10TAUR-Lab /Taur_CoT_Analysis_Project___microsoft__Phi-3-small-8k-instructtext10K<n<100K0 likes923 downloads2y agoHugging Face11TAUR-Lab /Taur_CoT_Analysis_Project___google__gemini-1.5-flash-001text10K<n<100K0 likes893 downloads2y agoHugging Face12patched-codes /static-analysis-evalA dataset of 76 Python programs taken from real Python open source projects (top 100 on GitHub), where each program is a file that has exactly 1 vulnerability as detected by a particular static analyzer (Semgrep), used in the paper Patched MOA: optimizing inference for diverse software development tasks. OpenAI used the synth-vuln-fixes and fine-tuned a new version of gpt-4o is now the SOTA on this benchmark. More details and code is available from their repo. More details on the benchmark… See the full description on the dataset page: https://huggingface.co/datasets/patched-codes/static-analysis-eval.textn<1K20 likes735 downloads1y agoHugging Face13winvoker /turkish-sentiment-analysis-dataset Dataset This dataset contains positive , negative and notr sentences from several data sources given in the references. In the most sentiment models , there are only two labels; positive and negative. However , user input can be totally notr sentence. For such cases there were no data I could find. Therefore I created this dataset with 3 class. Positive and negative sentences are listed below. Notr examples are extraced from turkish wiki dump. In addition, added some random text… See the full description on the dataset page: https://huggingface.co/datasets/winvoker/turkish-sentiment-analysis-dataset.texttext-classification100K<n<1M49 likes693 downloads3y agoHugging Face14ramankamran /retina-age-analysis Retina Age Analysis Dataset Dataset Description This dataset contains 9,857 retinal fundus images from 5,393 patients for age prediction tasks. Dataset Summary Task: Age prediction from retinal fundus images Images: 9,857 high-quality retinal images Patients: 5,393 unique patients Age Range: 5-97 years Image Format: JPEG Average Image Size: ~1 MB Supported Tasks Regression: Predict continuous age (5-97 years) Classification: Predict age group (5… See the full description on the dataset page: https://huggingface.co/datasets/ramankamran/retina-age-analysis.imageimage-classification1K<n<10K0 likes596 downloads11mo agoHugging Face15TAUR-Lab /Taur_CoT_Analysis_Project___mistralai__Mistral-7B-Instruct-v0.3text100K<n<1M0 likes594 downloads2y agoHugging Face16hugginglearners /amazon-reviews-sentiment-analysis Dataset Card for amazon reviews for sentiment analysis Dataset Summary One of the most important problems in e-commerce is the correct calculation of the points given to after-sales products. The solution to this problem is to provide greater customer satisfaction for the e-commerce site, product prominence for sellers, and a seamless shopping experience for buyers. Another problem is the correct ordering of the comments given to the products. The prominence of misleading… See the full description on the dataset page: https://huggingface.co/datasets/hugginglearners/amazon-reviews-sentiment-analysis.tabular1K<n<10K5 likes588 downloads4y agoHugging Face17TAUR-Lab /Taur_CoT_Analysis_Project___google__gemini-1.5-pro-001text10K<n<100K1 likes558 downloads2y agoHugging Face18ParsiAI /snappfood-sentiment-analysistexttext-classification10K<n<100K7 likes531 downloads2y agoHugging Face19ParsiAI /digikala-sentiment-analysistabulartext-classification1K<n<10K3 likes517 downloads2y agoHugging Face20NuBerea /source-analysisgated NuBerea Source Analysis Source-critical analysis of the Hebrew Bible, Septuagint, New Testament, Vulgate, and Second Temple literature. The dataset carries machine-generated source and tradition annotations at the verse level — the classical concerns of source criticism (documentary strata in the Old Testament, corpus structure in the New Testament, the pathway of Old Testament traditions into New Testament citation) expressed as structured data — together with semantic-domain… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/source-analysis.tabularfeature-extraction100K<n<1M0 likes517 downloads2d agoHugging Face21timchen0618 /browsecomp-plus-selected-tools-analysis-v1 BrowseComp-Plus: Selected Tools Analysis Side-by-side view of selected tool calls from a reference trajectory alongside the new agent trajectory conditioned on those steps. Retrieval model: Qwen3-Embedding-8BAgent model: gpt-oss-120bRun: traj_summary_ext_selected_tools_gpt-oss-120b_seed0 Columns Column Description query_id Query identifier rationale GPT rationale for why these k steps were selected from the reference trajectory selected_indices Step indices… See the full description on the dataset page: https://huggingface.co/datasets/timchen0618/browsecomp-plus-selected-tools-analysis-v1.tabularn<1K0 likes511 downloads6mo agoHugging Face22TAUR-Lab /Taur_CoT_Analysis_Project___Qwen__Qwen2-72B-Instructtext10K<n<100K0 likes480 downloads2y agoHugging Face23TAUR-Lab /Taur_CoT_Analysis_Project___claude-3-5-sonnet-20240620text10K<n<100K1 likes473 downloads2y agoHugging Face24hf-azure-internal /trending-models-analysishttps://github.com/pagezyhf/azure-cron/blob/main/trending_models_analysis.py text10K<n<100K3 likes468 downloads20h agoHugging Face25parallel-reasoner /Step-analysistext100K<n<1M0 likes465 downloads2mo agoHugging Face26NuBerea /translation-analysisgated NuBerea Translation Verse Texts Verse-level texts of historical Bible translations (Clementine Vulgate, Luther Bible 1545, Matthew's Bible 1537). Part of the NuBerea curated corpus estate of biblical and historical texts. Attribution Upstream Data Sources Source License Clementine Vulgate, NOCR Public Domain Luther Bible 1545, NOCR Public Domain Matthew's Bible 1537, Textus Receptus Bibles Public Domain NuBerea project. Licensed… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/translation-analysis.tabularfeature-extraction100K<n<1M0 likes458 downloads10d agoHugging Face27NuBerea /pseudepigrapha-analysisgated NuBerea Pseudepigrapha Analysis Derived linguistic datasets over pseudepigraphal literature, part of the NuBerea curated corpus estate. Covers the Greek and Latin witnesses of these texts along with a multilingual view across the available witness languages. License CC BY 4.0. Attribution Source Link License NuBerea project https://huggingface.co/NuBerea CC BY 4.0 tabularfeature-extraction10K<n<100K0 likes437 downloads2d agoHugging Face28Sp1786 /multiclass-sentiment-analysis-dataset Dataset Card for Dataset Name Dataset Summary This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Sp1786/multiclass-sentiment-analysis-dataset.tabulartext-classification10K<n<100K29 likes432 downloads3y agoHugging Face29NuBerea /lxx-analysisgated NuBerea Research: LXX Translation-Technique Noise Model Quantitative study of Septuagint translation technique: verse-by-verse measurements of where the ancient Greek translation (LXX) diverges from the Hebrew Masoretic Text, with book-level statistical summaries. The material lets researchers distinguish a translator's habitual working style — free versus literal rendering — from genuine textual anomalies worth close scholarly attention, putting on a measurable footing what LXX… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/lxx-analysis.tabularfeature-extraction10K<n<100K0 likes431 downloads2d agoHugging Face30NuBerea /septuagint-analysisgated NuBerea Septuagint Textual Analysis Curated datasets for study of the Septuagint (the ancient Greek translation of the Hebrew Bible), part of the NuBerea corpus estate of biblical and patristic texts. It gathers Septuagint verse texts, apparatus notes, and edition-comparison material into a set of ready-to-load configurations. Attribution This dataset derives from the following upstream sources, which require attribution: Source License Rahlfs… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/septuagint-analysis.tabularfeature-extraction10K<n<100K0 likes429 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.