datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AssetOpsBench
AssetOpsBench
AssetOpsBench is a specialized benchmark designed for evaluating Large Language Models (LLMs) and Multi-Agent systems in industrial operations. It focuses on the intersection of sensor data interpretation, maintenance logic, and Prognostics and Health Management (PHM).
The benchmark enables researchers to test how effectively AI agents can manage complex industrial assets, such as compressors and hydraulic pumps, by applying rule-based logic and diagnostic… See the full description on the dataset page: https://huggingface.co/datasets/ibm-research/AssetOpsBench.nihaisha-rag-assets
Nihaisha RAG Runtime Assets
Public production runtime assets for the nihaisha-rag-prototype project.
Scope and provenance
This repository contains generated RAG runtime assets, not source PDF files.
The corpus has 23 documents: 10 course-primary documents, 1 classic-primary candidate, and 12 related-reference documents.
Related-reference documents are 关联参考资料(非倪海厦著作). They must remain visibly separated from course-primary evidence.
The classic-primary candidate… See the full description on the dataset page: https://huggingface.co/datasets/JuneYao/nihaisha-rag-assets.AssetOpsBench
AssetOpsBench
AssetOpsBench is a specialized benchmark designed for evaluating Large Language Models (LLMs) and Multi-Agent systems in industrial operations. It focuses on the intersection of sensor data interpretation, maintenance logic, and Prognostics and Health Management (PHM).
The benchmark enables researchers to test how effectively AI agents can manage complex industrial assets, such as compressors and hydraulic pumps, by applying rule-based logic and diagnostic reasoning.… See the full description on the dataset page: https://huggingface.co/datasets/zhansingsong/AssetOpsBench.CoRe
CoRe: Benchmarking LLMs’ Code Reasoning Capabilities through Static Analysis Tasks
This repository hosts the CoRe benchmark, designed to evaluate the reasoning capabilities of large language models on program analysis tasks including data dependency, control dependency, and information flow. Each task instance is represented as a structured JSON object with detailed metadata for evaluation and reproduction.
It contains 25k data points (last update: Sep. 24th, 2025).
Each example is… See the full description on the dataset page: https://huggingface.co/datasets/lt-asset/CoRe.PhysBench-assets
PhysBench
🌐 Homepage | 🤗 Dataset | 📑 Paper | 💻 Code | 🔺 EvalAI
This repo contains additional assets for the test splits, as mentioned in the paper:
auxiliary_image.zip: The auxiliary image for simulation data. 🖼️
config.zip: The configuration files for simulation data. ⚙️
Happy experimenting! 😄
Other links:
PhysBench-test
PhysBench-train
PhysBench-media
alternative-asset-literacy-glossary
Alternative Asset Literacy Glossary
351 research-sourced financial terms across 6 categories: Alternative Assets, Art, DeFi & Crypto, ESG & Climate, Behavioral Economics, and Gender Lens Investing.
Built for the Alternative Asset Literacy iOS platform by Victoria Lee Case, Untitled_ LuxPerpetua Technologies, Inc.
Dataset Description
This glossary covers financial terminology used in alternative asset education. It is distinct from standard financial glossaries in… See the full description on the dataset page: https://huggingface.co/datasets/UntitledFinancial/alternative-asset-literacy-glossary.
