CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ibm-research /AssetOpsBench AssetOpsBench AssetOpsBench is a specialized benchmark designed for evaluating Large Language Models (LLMs) and Multi-Agent systems in industrial operations. It focuses on the intersection of sensor data interpretation, maintenance logic, and Prognostics and Health Management (PHM). The benchmark enables researchers to test how effectively AI agents can manage complex industrial assets, such as compressors and hydraulic pumps, by applying rule-based logic and diagnostic… See the full description on the dataset page: https://huggingface.co/datasets/ibm-research/AssetOpsBench.textquestion-answeringn<1K46 likes822 downloads4mo agoHugging Face02JuneYao /nihaisha-rag-assets Nihaisha RAG Runtime Assets Public production runtime assets for the nihaisha-rag-prototype project. Scope and provenance This repository contains generated RAG runtime assets, not source PDF files. The corpus has 23 documents: 10 course-primary documents, 1 classic-primary candidate, and 12 related-reference documents. Related-reference documents are 关联参考资料(非倪海厦著作). They must remain visibly separated from course-primary evidence. The classic-primary candidate… See the full description on the dataset page: https://huggingface.co/datasets/JuneYao/nihaisha-rag-assets.question-answering1 likes197 downloads2mo agoHugging Face03zhansingsong /AssetOpsBench AssetOpsBench AssetOpsBench is a specialized benchmark designed for evaluating Large Language Models (LLMs) and Multi-Agent systems in industrial operations. It focuses on the intersection of sensor data interpretation, maintenance logic, and Prognostics and Health Management (PHM). The benchmark enables researchers to test how effectively AI agents can manage complex industrial assets, such as compressors and hydraulic pumps, by applying rule-based logic and diagnostic reasoning.… See the full description on the dataset page: https://huggingface.co/datasets/zhansingsong/AssetOpsBench.textquestion-answeringn<1K1 likes83 downloads4mo agoHugging Face04lt-asset /CoRe CoRe: Benchmarking LLMs’ Code Reasoning Capabilities through Static Analysis Tasks This repository hosts the CoRe benchmark, designed to evaluate the reasoning capabilities of large language models on program analysis tasks including data dependency, control dependency, and information flow. Each task instance is represented as a structured JSON object with detailed metadata for evaluation and reproduction. It contains 25k data points (last update: Sep. 24th, 2025). Each example is… See the full description on the dataset page: https://huggingface.co/datasets/lt-asset/CoRe.question-answering10K<n<100K1 likes52 downloads1y agoHugging Face05WeiChow /PhysBench-assets PhysBench 🌐 Homepage | 🤗 Dataset | 📑 Paper | 💻 Code | 🔺 EvalAI This repo contains additional assets for the test splits, as mentioned in the paper: auxiliary_image.zip: The auxiliary image for simulation data. 🖼️ config.zip: The configuration files for simulation data. ⚙️ Happy experimenting! 😄 Other links: PhysBench-test PhysBench-train PhysBench-media question-answering1K<n<10K0 likes43 downloads2y agoHugging Face06UntitledFinancial /alternative-asset-literacy-glossary Alternative Asset Literacy Glossary 351 research-sourced financial terms across 6 categories: Alternative Assets, Art, DeFi & Crypto, ESG & Climate, Behavioral Economics, and Gender Lens Investing. Built for the Alternative Asset Literacy iOS platform by Victoria Lee Case, Untitled_ LuxPerpetua Technologies, Inc. Dataset Description This glossary covers financial terminology used in alternative asset education. It is distinct from standard financial glossaries in… See the full description on the dataset page: https://huggingface.co/datasets/UntitledFinancial/alternative-asset-literacy-glossary.text-classificationn<1K0 likes19 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.