datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
MetricScenes
MetricScenes
A metrically-grounded, in-the-wild dataset. For more details, please visit the project page.
Paper
Title: Honey, I Shrunk the Arc de Triomphe!Authors: Yuanbo Xiangli, Hanyu Chen, Xueqing Tsang, Noah SnavelyProject page: https://metricscenes.github.io/
Abstract
Metric scale monocular geometry estimation has seen significant progress through large-scale data aggregation, yet current foundation models suffer from a persistent… See the full description on the dataset page: https://huggingface.co/datasets/yx642/MetricScenes.Qwen3.8-27B-GGUF-metrics
Qwen3.8-27B GGUF, everything behind the numbers
This is the working record for
AtomicChat/Qwen3.8-27B-GGUF.
Every figure in that model card came from a file in here, including the ones
about other publishers' builds.
The point of publishing it is simple. A quantization comparison is only worth
reading if someone else can run it, and that needs three things nobody usually
ships: the exact reference the numbers were measured against, the exact text
they were measured on, and the… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/Qwen3.8-27B-GGUF-metrics.llmproj-training-metricsbokeh-eval-metricastokenizers-dependents
tokenizers metrics
This dataset contains metrics about the huggingface/tokenizers package.
Number of repositories in the dataset: 11460
Number of packages in the dataset: 124
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 14 packages that have more than 1000 stars.
There are 41… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/tokenizers-dependents.language-metric-data# This dataset contains the entire content of three files loaded as a single example:
# - `languages_list.pkl`: A pickled list of language strings.
# - `average_distances_matrix.npy`: A NumPy matrix converted to a list of lists of floats.
# - `distances_matrices.pkl`: A pickled dict of dicts of NumPy matrices.
# It is converted into a list of records where each record corresponds to a dataset with a nested list of models and their associated distance matrices.
#atomic-metrics-experiments
Atomic Metrics Experiment Artifacts
Run directories from
Atomic Metrics, including
extracted metric banks, generated domain prompts, batch/refine snapshots, and
BT/LR eval summaries.
Logs are omitted. API keys are not included; scoring used environment
credentials at runtime.
pythia-training-metrics Dataset for storing training metrics of pythia modelstransformers-dependents
transformers metrics
This dataset contains metrics about the huggingface/transformers package.
Number of repositories in the dataset: 27067
Number of packages in the dataset: 823
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 65 packages that have more than 1000 stars.
There are 140… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/transformers-dependents.navsim-metric-caches-from-a100pythia-training-metrics-40m-qk-layernorm Dataset for storing training metrics of pythia modelsMuse-Glimmer-30B-GGUF-metrics
Muse Glimmer 30B GGUF — raw metrics
Every log behind the numbers in
AtomicChat/Muse-Glimmer-30B-GGUF.
Published unfiltered, so any figure in the model card can be checked or disputed.
Layout
Path
Contents
kld/
llama-perplexity --kl-divergence output, per build and per corpus
bench/
llama-bench -o json
speculative/
llama-server logs with and without the drafter
layouts/
per-tensor type map of every GGUF
conversion/
convert_hf_to_gguf.py logs… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/Muse-Glimmer-30B-GGUF-metrics.sparse-metric-anchors-ycb
Sparse Metric Anchors — YCB benchmark
The 42-object benchmark behind the paper Sparse Metric Anchors for a Single-View 3D
Generative Prior: The Output Frame Is the Bottleneck (ISIR, Sorbonne Université,
2026). Code and paper: github.com/635jack/sparse-metric-anchors — its colab/reproduce.ipynb recomputes every table of the paper from this dataset on a CPU runtime.
The paper asks what limits the injection of a few metric measurements — tactile
contacts, one depth map — into a… See the full description on the dataset page: https://huggingface.co/datasets/jack635/sparse-metric-anchors-ycb.gradio-dependents
Dataset Card for "gradio-dependents"
More Information needed
DeepSeek-V4.1-Flash-NVFP4-metrics
DeepSeek-V4.1-Flash-NVFP4 metrics
Everything behind the numbers in AtomicChat/DeepSeek-V4.1-Flash-NVFP4-nvidia.
logprobs/lp-<run>-<corpus>.npz: the raw top-512 log probabilities of every measurement run, 49,152 scored
positions each: ref, ref-repeat, ref-r3, ref-b1 (batch size 1) for the original; flat, flat-r2,
flat-r3 for the uncalibrated cast; nvidia, nvidia-r2, nvidia-r3 for the calibrated checkpoint.
logs/kld-<run>-<corpus>.json: the KL lower bound per run against ref… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/DeepSeek-V4.1-Flash-NVFP4-metrics.Ornith-1.5-35B-A3B-GGUF-metricsdatasets-dependents
datasets metrics
This dataset contains metrics about the huggingface/datasets package.
Number of repositories in the dataset: 4997
Number of packages in the dataset: 215
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 22 packages that have more than 1000 stars.
There are 43… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/datasets-dependents.omnidocbench-render-compare
OmniDocBench Render-and-Compare
This dataset contains the rendered HTML reconstructions and comparison images produced
by a render-and-compare pipeline — a reference-free visual similarity evaluation
framework for OCR systems.
Overview
The pipeline processes each page of OmniDocBench through
a Qwen3.5-122B-A10B OCR model, renders the structured output back to a PNG via HTML
(reconstructed.png), and compares it against the original page scan (masked_original.png)
using… See the full description on the dataset page: https://huggingface.co/datasets/gt-free-ocr-metrics/omnidocbench-render-compare.evaluate-dependents
evaluate metrics
This dataset contains metrics about the huggingface/evaluate package.
Number of repositories in the dataset: 106
Number of packages in the dataset: 3
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 1 packages that have more than 1000 stars.
There are 2 repositories… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/evaluate-dependents.accelerate-dependents
accelerate metrics
This dataset contains metrics about the huggingface/accelerate package.
Number of repositories in the dataset: 727
Number of packages in the dataset: 37
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 10 packages that have more than 1000 stars.
There are 16… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/accelerate-dependents.llm-metric-tulumetric-vqa-qwen3-235b-bundle
Qwen3-VL-235B bundle (v2) — metric_vqa probes for RITS
Self-contained bundle for running Qwen3-VL-235B on every metric_vqa probe.
Built 2026-05-17.
Setup
pip install requests pandas Pillow tqdm
export RITS_API_BASE="https://.../v1"
export RITS_API_KEY="..."
Run a single condition
python scripts/run_qwen235b.py \
--manifest manifests/in_the_wild_vision.csv \
--output-dir out/in_the_wild_vision
Output: predictions.csv + metrics.json in the output dir.… See the full description on the dataset page: https://huggingface.co/datasets/bmeivar/metric-vqa-qwen3-235b-bundle.metric-mamba-ml2021-hungyi-corpus
Dataset Card for "metric-mamba-ml2021-hungyi-corpus"
More Information needed
diffusers-dependents
diffusers metrics
This dataset contains metrics about the huggingface/diffusers package.
Number of repositories in the dataset: 160
Number of packages in the dataset: 2
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 0 packages that have more than 1000 stars.
There are 3 repositories… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/diffusers-dependents.pytorch-image-models-dependents
pytorch-image-models metrics
This dataset contains metrics about the huggingface/pytorch-image-models package.
Number of repositories in the dataset: 3615
Number of packages in the dataset: 89
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 18 packages that have more than 1000… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/pytorch-image-models-dependents.optimum-dependents
optimum metrics
This dataset contains metrics about the huggingface/optimum package.
Number of repositories in the dataset: 19
Number of packages in the dataset: 6
Package dependents
This contains the data available in the used-by
tab on GitHub.
Package & Repository star count
This section shows the package and repository star count, individually.
Package
Repository
There are 0 packages that have more than 1000 stars.
There are 0 repositories that… See the full description on the dataset page: https://huggingface.co/datasets/open-source-metrics/optimum-dependents.Ling-3.0-flash-GGUF-metrics
Ling-3.0-flash — quantization metrics
Everything measured while building the GGUF line for inclusionAI/Ling-3.0-flash: raw logs, per-rung numbers and the importance matrix statistics. Published so the quant table can be checked rather than trusted.
Quants live in AtomicChat/Ling-3.0-flash-GGUF.
Layout
metrics/
grid-table.json per rung: size, bpw, mean/99% KLD, top-1 agreement
kld-results.json raw parser output of every KL divergence run… See the full description on the dataset page: https://huggingface.co/datasets/AtomicChat/Ling-3.0-flash-GGUF-metrics.Metric-Bench
Metric-Bench Test
Metric-Bench Test set evaluates metric spatial understanding from indoor RGB images and explicit anchor measurements. Each example asks for a physical measurement of a referred object or the distance between two objects. Inputs consist of an image and an English question; the reference answer is a numeric JSON object.
This release contains 1,340 questions, 134 images and 20 scenes, with one test split. It is a reconstructed test revision and does not reproduce… See the full description on the dataset page: https://huggingface.co/datasets/yulingxi/Metric-Bench.cleangov-local-fiscal-metrics
지방재정365 통합공시·재정분석 Open API (수의계약·행사축제·업무추진비·의회경비·채무·기금·지방보조금·재정분석 결과)
지방재정365 재정데이터개방 허브의 "지방재정 통합공시" 분류 77 서비스와 "성과/평가 > 재정분석결과" 5 서비스, 합계 82 서비스. 자치단체별로 공시하는 지표군 (수의계약비율, 행사·축제경비 비율·편성내역·원가회계, 업무추진비 비율·절감률·기관운영·시책추진, 지방의회 관련경비·국외여비, 채무·지방채·보증채무·채권, 기금현재액, 지방보조금(2016·2020·2021)·지방보조금비율, 재정자립도·재정자주도·통합재정수지(결산·최종), 사회보장적 수혜금, 예비비, 통장이장반장 보상금, 공무원 관련경비, 지방교부세 인센티브·자체노력 반영, 재정운용계획 등) 과 행정안전부 지방재정분석 결과 지표다. 동종 지자체 비교(비교군) 의 기준 자료이며 대부분 자치단체 × 회계연도 × 지표 단위의 집계값이라 사건 단위 연결에는 제한이 있다.… See the full description on the dataset page: https://huggingface.co/datasets/eddmpython/cleangov-local-fiscal-metrics.pip
Dataset Card for "pip"
More Information needed
