reproducibility
naacl22_causalDistilBERT_instance_1naacl22_causalDistilBERT_instance_3naacl22_causalDistilBERT_instance_2norms_establish_check_reproducibility_7norms_establish_check_reproducibility_11norms_establish_check_reproducibility_1norms_establish_check_reproducibility_3norms_establish_check_reproducibility_15
reproducibility-datakit-technical-reportSampled parquet for the gridfm-datakit technical report diversity plots.
How to reproduce: scripts/datakit_report/README.md on branch genco-paper-repro.
shapleymcg-qwen3-30b-a3b-reproducibility
ShapleyMCG Qwen3-30B-A3B reproducibility artifacts
This dataset preserves the calibration statistics, exact corrected-R10
EXL3/MCG candidates, source and corpus identities, BF16 teacher/student logits,
tokenwise KLD, allocations, attribution ledgers, hashes, and publication
receipts for the Qwen3-30B-A3B experiments in
brandonmmusic-max/shapleymcg.
The complete cross-checkpoint
results ledger
and
method specification
distinguish the predecessor routed-p2 allocator from the full… See the full description on the dataset page: https://huggingface.co/datasets/brandonmusic/shapleymcg-qwen3-30b-a3b-reproducibility.Squidiff_reproducibility
🦑 Squidiff Reproducibility (Code + Processed Dataset)
English | 简体中文
This repository is a comprehensive, ready-to-use replication bundle for Squidiff. It contains annotated replication Jupyter notebooks alongside all the heavily processed intermediate .h5ad matrices (approx 22.8 GB in total) required to seamlessly reproduce the figures and model results without wrestling with data wrangling.
Note: This repository is cloned and extended from the official Squidiff reproducibility… See the full description on the dataset page: https://huggingface.co/datasets/zyzhou110/Squidiff_reproducibility.reproducibility-powermodels-setup2Corrected PowerModels JSON for the GENCO setup-2 (from-disk) runtime experiments.
How to reproduce: scripts/runtime/README.md on branch genco-paper-repro.
GradeSQL-reproducibility-data
Data Summary
This repository contains the reproducibility data for GradeSQL, a framework for fine-tuning text-to-SQL ORM models. It provides all the data required to reproduce the experiments and fine-tuning setups described in the GradeSQL paper.
The data is intended for researchers and practitioners who want to train, evaluate, or reproduce results from GradeSQL.For details about the methodology, usage, and tutorials, please refer to the main project repository: GradeSQL GitHub.
atec2026-task-e-reproducibility
ATEC2026 L0 Task E 复现数据与日志
本仓库保存 Datawhale ATEC2026 线上赛 L0「桌面整理 Task E」赛后开源复现所需的大文件。
配套 GitHub 教程与代码:
https://github.com/datawhalechina/every-embodied/tree/main/15-Challenge%E7%AB%9E%E8%B5%9B/ATEC2026/L0-%E6%A1%8C%E9%9D%A2%E6%95%B4%E7%90%86TaskE
配套模型权重:
https://huggingface.co/Datawhale/atec2026-task-e-act-seed1-best
文件说明
data/final_100demos_filtered_split_100m/trajectory_filtered.hdf5.part-0000 ... part-0336
ACT seed1 best 方案训练使用的过滤后 100-demo HDF5,按 100MiB… See the full description on the dataset page: https://huggingface.co/datasets/Datawhale/atec2026-task-e-reproducibility.
