anchor
keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-Anchored-Tensorsllavaqwen3-1.7b-finetune-nm-mask-moe-sparse-4e-2k-4of8-imp-anchor_20260812_071057time-anchor-modernbert-32mGLM-4.1V-9B-Anchor-Windows-GGUFGLM-4.1V-9B-Anchor-Ubuntu-GGUFAnchorCoder_v3-GGUFDSV4-Flash-0731-ablit-anchored-GUFFAnchor-Classification-DMV
Datasets
All datasets matching “anchor”TrainingData_Stage3
AnchorSR Stage3 · metric-v1.0
直接选择 Small / Large
配置
训练题数
用途
small
1,000,000
先验证答案监督/先验恢复,按新版 Large 联合分布抽样
large
89,801,853
筛选后的完整训练集合,包含 Small 全部样本
from datasets import load_dataset
data = load_dataset('AnchorSR/TrainingData_Stage3', 'small', # 或 large
revision='metric-v1.0', streaming=True)
这是对 scaling-v1.0 的语义筛选与统一任务分类,不是增加新数据源。
Large 从 89,828,269 题保留 89,801,853 题,隔离 26,416 题。
旧标签 scaling-v1.0 / video-v1.0 / large-v1.0… See the full description on the dataset page: https://huggingface.co/datasets/AnchorSR/TrainingData_Stage3.anchoral-paper-artefactsArtefacts related to the paper AnchorAL: Computationally Efficient Active Learning for Large and Imbalanced Datasets (Lesci and Vlachos, 2024) published at the NAACL 2024 conference.
These artefacts can be reproduced using the code available at github.com/pietrolesci/anchoral.
The outputs/ folder includes the raw files created by the individual experiments.
The results/ folder contains the exported metrics and configurations that are used to complete the analysis and create the tables and… See the full description on the dataset page: https://huggingface.co/datasets/pietrolesci/anchoral-paper-artefacts.Anchor-Lab
Anchor-Lab Dataset Card
Dataset Description
Anchor-Lab is a tabular robotics dataset captured from Anchor Lab, a sim-to-real transfer laboratory from NVIDIA Robotics. The dataset is designed to support calibration of physics simulation against physical robot measurements for zero-shot sim-to-real deployment.
The public release contains long-form parquet tables of robot experiment telemetry. Each row records a timestamped scalar measurement for an experiment and… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Anchor-Lab.history-anchor-100-traces
History Anchor 100 — Model Trajectories
*Per-(model × condition × scenario set × seed) raw outputs from the paper "History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions".*
This dataset contains the full set of model decisions that back every figure and table in the paper. Use it to:
audit a single model's behaviour scenario-by-scenario,
recompute headline metrics without re-running the (paid) API sweeps,
mine reasoning_content traces from models that expose… See the full description on the dataset page: https://huggingface.co/datasets/albertoRodriguez97/history-anchor-100-traces.controlled_anchor_v1_support_switch
Controlled ICIL Anchor V1 Support Switch
LeRobot conversion of the Anchor V1 controlled ICIL collection.
Source HDF5:
/ibex/project/c2090/jian/icil_openpi/ICIL/data/manifest_collection_v1/controlled_anchor_v1_60a_6p_3j_6obj_res256_lzf_merged.hdf5
OpenPI sidecars are stored under meta/controlled_icil/.
app-rank-anchors
App Rank Anchors
Community-federated public app-store calibration anchors for the
AppScope open app-intelligence
stack.
Each row is a public fact — a segment + rank + observed download flow —
derived from the public Google Play realInstalls delta over a time window
paired with an app's chart rank in that window. Pooling these anchors across
self-hosting contributors lets the Garg–Telang download estimator calibrate
absolute scale (scale_b) per (platform, category, country)… See the full description on the dataset page: https://huggingface.co/datasets/Ahad690/app-rank-anchors.
