datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CostNav-Teleop-Dataset
CostNav Teleop Dataset
Dataset Summary
The CostNav Teleop Dataset is a large-scale collection of human teleoperation recordings for robot navigation in an urban sidewalk simulation environment. It was collected as part of the CostNav benchmark, which evaluates navigation systems using real-world economic cost and revenue metrics rather than purely technical metrics.
The dataset contains 2,203 teleoperation episodes totaling 50.2 hours of driving… See the full description on the dataset page: https://huggingface.co/datasets/maum-ai/CostNav-Teleop-Dataset.cruise-costs
Real All-In Cruise Costs — by CruiseClarify
📅 This is a dated snapshot — refreshed September 2026
Cruise prices change constantly. Every figure here is a point-in-time value, not
live data, accurate as of September 2026. The next quarterly refresh is due
December 2026 — after that, treat this copy as historical.
If you are reading or reusing this a quarter or more after September 2026, the
numbers will have drifted. Always pull the current version before quoting… See the full description on the dataset page: https://huggingface.co/datasets/CruiseClarify/cruise-costs.llm-cost-same-prompt
Measured per-call LLM cost — same prompt, every model
Vendors publish prices per million tokens. Nobody publishes what one call actually costs, because
that depends on how many tokens the model chooses to emit — and on the same question models differ by
more than an order of magnitude. One model finishes a JSON extraction in 23 tokens; another writes 300.
This dataset sends a fixed set of prompts to every model at temperature 0, every night, and records
the cost computed from… See the full description on the dataset page: https://huggingface.co/datasets/mario0369/llm-cost-same-prompt.fpga_cost_model_kernel_data
FPGA HLS Kernel Cost-Model Data
Evolved Vitis HLS C++ kernels paired with their ground-truth Vitis HLS
csynth results. Each row is one generated program from an evolutionary FPGA
optimisation run, linked to its kernel source, evaluator report.json, and raw
synthesis report.
Each row carries a split label: train marks the original benchmarks used
to fit the analytical cost model's learned correction term, and holdout marks
benchmarks added afterwards that were not used for… See the full description on the dataset page: https://huggingface.co/datasets/adimnaku/fpga_cost_model_kernel_data.fpga_cost_model_kernel_data_attention_p2
FPGA HLS Kernel Cost-Model Data
Evolved Vitis HLS C++ kernels paired with their ground-truth Vitis HLS
csynth results. Each row is one generated program from an evolutionary FPGA
optimisation run, linked to its kernel source, evaluator report.json, and raw
synthesis report.
Each row carries a split label: train marks the original benchmarks used
to fit the analytical cost model's learned correction term, and holdout marks
benchmarks added afterwards that were not used for… See the full description on the dataset page: https://huggingface.co/datasets/adimnaku/fpga_cost_model_kernel_data_attention_p2.SWEbench-Verified-eval150-u355-M2.7-orch-cost-20260923
Fixed Solo350 u355 + MiniMax-M2.7: orchestration cost study
Best observed cost tradeoff: compact coordinator decisions plus soft review at the existing hard limit (at most 12 worker turns). M2.7 metered token cost falls 59.3%, while mean solved tasks decrease from 90.00 to 87.67/150. Accuracy equivalence was not established.
This closed study contains 6 designs and 16 complete independent runs on the same 150 tasks (2400 scored task/run pairs), each with an independent audit.… See the full description on the dataset page: https://huggingface.co/datasets/CharlieLLL/SWEbench-Verified-eval150-u355-M2.7-orch-cost-20260923.457k-prices-build-a-burger
457,352 prices: 10 categories, 12 U.S. ZIPs, 29 days
Burger Ingredient Prices Raw Dataset (2026)
How do listed and package-standardized prices for common burger components differ across U.S. ZIP markets and days?
This fixed research snapshot contains 457,352 unaggregated, quality-filtered price observations across 10 burger-component categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV preserves product titles… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/457k-prices-build-a-burger.gcp-cloud-billing-costbusiness-cost-matrix-scenario-suite
The Business Cost-Matrix Scenario Suite
Twenty real-shaped business decision problems, each with a cost matrix built from public
evidence instead of an assumed symmetric loss.
Cost-sensitive learning needs a cost matrix, and in practice that matrix is almost always invented.
Someone picks 5:1 or 10:1, and the experiment then measures the behaviour of the assumption rather
than the behaviour of the world. This dataset exists so nobody has to keep doing that.
Every non-zero cost… See the full description on the dataset page: https://huggingface.co/datasets/Nafeel123/business-cost-matrix-scenario-suite.EMNLP_Cost-Aware-Protocol-Routing
Cost-Aware Protocol Routing: Matched Protocol Outcomes
The short version. We ran the same 6,803 reasoning problems through four
different LLM collaboration setups — from a single direct answer up to a
four-agent deliberation — and recorded, for every problem, which ones got it
right. Then we asked whether a model can look at a problem beforehand and
predict which setup is worth paying for.
It can predict whether it will fail. It cannot predict which collaboration
protocol will… See the full description on the dataset page: https://huggingface.co/datasets/AgentsSci/EMNLP_Cost-Aware-Protocol-Routing.asia-health-cost-2024-consolidated
Asia Health Cost 2024 — Consolidated
Consolidated 2024 fiscal-year medical operations & cost dataset for a pan-Asia healthcare enterprise
(China / Japan / India). Created by merging three country-level, de-identified source datasets and
normalising every cost to USD.
Namespace note: the task referenced the source/output under the medi-core namespace, which is not
accessible with the current credentials. The identical pipeline was executed under the toolathon123
namespace:… See the full description on the dataset page: https://huggingface.co/datasets/toolathon123/asia-health-cost-2024-consolidated.193k-prices-period-care-atlas
192,500 prices: 9 categories, 12 U.S. ZIPs, 29 days
Period Care Prices Raw Dataset (2026)
How do listed and package-standardized prices for reusable and disposable period-care products vary across U.S. ZIP markets and days?
This fixed research snapshot contains 192,500 unaggregated, quality-filtered price observations across 9 period-care categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV preserves product… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/193k-prices-period-care-atlas.ddr5-ram-prices-raw-dataset-2026
19,705 raw U.S. DDR5 RAM price observations across 12 ZIP markets and 29 days.
DDR5 RAM Prices Raw Dataset (2026)
Analyze 19,705 unaggregated product-level listed retail prices for standalone DDR5 memory modules and homogeneous kits across 12 U.S. ZIP markets from July 13 through August 10, 2026. The single analysis-ready CSV preserves titles, dates, geography, package quantities, listed prices, and a source-neutral comparable-price field.
What “raw” means here: unaggregated… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/ddr5-ram-prices-raw-dataset-2026.373k-prices-condiment-economy
372,714 prices: 5 categories, 12 U.S. ZIPs, 29 days
Condiment Prices Raw Dataset (2026)
How do listed and package-standardized prices for pickles, mayonnaise, ketchup, mustard, and pickle relish vary across U.S. ZIP markets and days?
This fixed research snapshot contains 372,714 unaggregated, quality-filtered price observations across 5 condiment categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV preserves… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/373k-prices-condiment-economy.60-5k-prices-stay-cool
60,453 prices: 6 categories, 12 U.S. ZIPs, 29 days
Air Conditioner Prices Raw Dataset (2026)
How do listed prices for portable, window, through-wall, and mini-split air conditioners vary across U.S. ZIP markets and days?
This fixed research snapshot contains 60,453 unaggregated, quality-filtered price observations across 6 residential air-conditioning categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/60-5k-prices-stay-cool.grocery-prices-raw-dataset-2026
1,207,833 prices: 17 categories, 12 U.S. ZIPs, 29 days
Grocery Prices Raw Dataset (2026)
How do listed and package-standardized grocery prices vary across 17 food and beverage categories, 12 selected U.S. ZIP markets, and 29 days?
This fixed research snapshot contains 1,207,833 unaggregated, quality-filtered price observations across 17 grocery categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV preserves… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/grocery-prices-raw-dataset-2026.313k-prices-nyc-vs-la-57-categories
313,126 prices: 57 categories, 2 U.S. ZIPs, 29 days
NYC vs LA Retail Prices Raw Dataset (2026)
How do listed and package-standardized prices vary between selected New York and Los Angeles ZIP markets across 57 everyday categories and 29 days?
This fixed research snapshot contains 313,126 unaggregated, quality-filtered price observations across 57 categories, 2 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV preserves… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/313k-prices-nyc-vs-la-57-categories.vn-provinces-spatial-living-cost-index
Vietnam provinces spatial living cost index (Ha Noi = 100)
Spatial living-cost (sinh hoạt) price index by province relative to Ha Noi (= 100). Coverage 2011-2024. Provinces only (no regional or national aggregate in the source table). Geographic labels are English (UN/GSO style ASCII romanization). Province names follow ar_core.vn_geo (historical 63-province system).
Figures
Hero
Comparison
Color key
Files
provinces (882 rows)
data/provinces.csv… See the full description on the dataset page: https://huggingface.co/datasets/letrinhan/vn-provinces-spatial-living-cost-index.wellness-supplement-prices-raw-dataset-2026
53,198 prices: 5 categories, 12 U.S. ZIPs, 29 days
Wellness Supplement Prices Raw Dataset (2026)
How do listed and package-standardized prices vary across five everyday wellness supplement categories, 12 selected U.S. ZIP markets, and 29 days?
This fixed research snapshot contains 53,198 unaggregated, quality-filtered price observations across 5 categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV preserves… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/wellness-supplement-prices-raw-dataset-2026.house-cost-prediction-multivariances
🏡 House Cost Prediction by Multi-Variances
A comprehensive dataset designed for real estate price prediction, data science experimentation, and machine learning model benchmarking.This dataset simulates global property listings with realistic variations in city, area, price, and socioeconomic indicators.
📘 Overview
The House Cost Prediction by Multi-Variances dataset provides one million synthetic yet statistically realistic property listings.Each entry contains… See the full description on the dataset page: https://huggingface.co/datasets/bdstar/house-cost-prediction-multivariances.fpga_cost_model_kernel_data_mamba_p2
FPGA HLS Kernel Cost-Model Data
Evolved Vitis HLS C++ kernels paired with their ground-truth Vitis HLS
csynth results. Each row is one generated program from an evolutionary FPGA
optimisation run, linked to its kernel source, evaluator report.json, and raw
synthesis report.
Each row carries a split label: train marks the original benchmarks used
to fit the analytical cost model's learned correction term, and holdout marks
benchmarks added afterwards that were not used for… See the full description on the dataset page: https://huggingface.co/datasets/adimnaku/fpga_cost_model_kernel_data_mamba_p2.us-tax-cost-of-living-relocation
US state and local tax, cost of living, and relocation comparisons (2026 tax year)
Three datasets behind estimatetax.net, published so the figures
can be checked, reused and cited rather than taken on trust.
Files
condados.csv — 3132 county effective property tax rates
The effective property tax rate for every US county and county-equivalent with a published
rate, derived from the US Census Bureau's American Community Survey. division_type records… See the full description on the dataset page: https://huggingface.co/datasets/estimatetax/us-tax-cost-of-living-relocation.ai-development-cost-benchmark
AI Development Cost Benchmark 2026
How much does AI development cost in 2026? This open dataset provides structured cost benchmarks for 8 categories of AI development projects across 3 complexity tiers, with 24 records covering cost ranges, timelines, team sizes, deliverables, and tech stacks.
Published and maintained by Salt Technologies AI, the AI engineering division of Salt Technologies (14+ years, 800+ projects delivered).
Quick Links
Interactive dataset page:… See the full description on the dataset page: https://huggingface.co/datasets/salttechno/ai-development-cost-benchmark.nvme-ssd-prices-raw-dataset-2026
108,196 raw U.S. NVMe SSD price observations across 12 ZIP markets and 29 days.
NVMe SSD Prices Raw Dataset (2026)
Analyze 108,196 unaggregated product-level listed retail prices for NVMe solid-state drives across 12 U.S. ZIP markets from July 13 through August 10, 2026. The single analysis-ready CSV preserves titles, dates, geography, package quantities, listed prices, and a source-neutral comparable-price field.
What “raw” means here: unaggregated product-level… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/nvme-ssd-prices-raw-dataset-2026.home-care-cost-model
Home Care Cost Model — Datasets
Open data companion to the Home Care Cost Model reference framework. Eight
CSV datasets totalling roughly 9,200 rows span three classes: raw source
retrievals, hand-curated reference tables, and engine-derived scenario
grids. All values are denominated in Canadian dollars (CAD) and indexed to
the 2026 taxation year.
Dataset classes
Class 1 — Sources (raw, credited)
Stored under sources/<organisation>/. Every file has a sibling… See the full description on the dataset page: https://huggingface.co/datasets/davecook1985/home-care-cost-model.wet-vs-dry-pet-food-prices-raw-dataset-2026
44,583 prices: 4 categories, 12 U.S. ZIPs, 29 days
Wet vs Dry Pet Food Prices Raw Dataset (2026)
How do listed and package-standardized prices compare across wet and dry dog and cat food, 12 selected U.S. ZIP markets, and 29 days?
This fixed research snapshot contains 44,583 unaggregated, quality-filtered price observations across 4 categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026. The analysis-ready CSV preserves product titles… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/wet-vs-dry-pet-food-prices-raw-dataset-2026.onion-prices-raw-dataset-2026
17,274 raw U.S. onion price observations across 12 ZIP markets and 29 days.
Onion Prices Raw Dataset (2026)
Analyze 17,274 unaggregated product-level listed retail prices for fresh common onions across 12 U.S. ZIP markets from July 13 through August 10, 2026. The single analysis-ready CSV preserves titles, dates, geography, package quantities, listed prices, and a source-neutral comparable-price field.
What “raw” means here: unaggregated product-level observations after… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/onion-prices-raw-dataset-2026.coffee-pod-capsule-prices-raw-dataset-2026
Coffee Pod & Capsule Prices Raw Dataset (2026)
Analyze 151,252 product-level listed retail prices for single-serve coffee pods and capsules across 12 U.S. ZIP markets and every date from July 13 through August 10, 2026. The CSV preserves full product titles, listed package prices, resolved pod counts, and a comparable 24-pod package-equivalent price.
What “raw” means here: unaggregated, product-level retail price observations after quality filtering. The file also contains… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/coffee-pod-capsule-prices-raw-dataset-2026.119k-prices-petflation-2026
119,316 prices: 11 categories, 12 U.S. ZIPs, 29 days
119K Prices: Petflation 2026
How do pet food, treats, litter, grooming, cleanup supplies, and training-pad prices compare across U.S. ZIP markets when package quantities are standardized within each category?
This fixed research snapshot contains 119,316 unaggregated, quality-filtered retail price observations across 11 pet-care categories, 12 U.S. ZIP markets, and 29 consecutive dates from July 21 through August 18, 2026.… See the full description on the dataset page: https://huggingface.co/datasets/costinflation/119k-prices-petflation-2026.costco_long_practice
