RU
Models
All models matching “RU”Datasets
All datasets matching “RU”transformers_circleci_workflow_runsinfini-news-corpus
INFINI-NEWS Corpus
🔎 Search this corpus online: query it with sub-second full-text search and n-gram counts — in the browser or via a public, keyless REST API, no download required — at infini-news.uni-graz.at (API reference).
A multilingual news corpus extracted from
Common Crawl CC-News WARC files.
One row per article, with body text extracted via
trafilatura,
WARC provenance, and derived metadata (publish date, language, topic,
byte hashes) in a single flat schema. Covers… See the full description on the dataset page: https://huggingface.co/datasets/ruggsea/infini-news-corpus.cherrl-runsulvr_subset
ULVR stage-0 subsets (latent + source)
Curated, nested subsets of the Unified Visual Latent Reasoning (ULVR) stage-0
training data. Each subset folder is self-contained and ships both:
latent/ — pre-computed teacher latents, identical schema to
RuoliuYang/step0-all
source/ — the matching source samples (images + question/answer +
messages), identical schema to
RuoliuYang/ULVR_v2_clean
Latents and source rows are joinable by sample_id (within a category).
Folder… See the full description on the dataset page: https://huggingface.co/datasets/RuoliuYang/ulvr_subset.ditec-wdn--
Dataset Card for DiTEC-WDN
Dataset Summary
DiTEC-WDN Dataset consists of 36 Water Distribution Networks (WDNs). Each network has unique 1,000 scenarios with distinct characteristics.
Scenario represents a timeseries of directed shared-topology graphs, referred to as states or snapshots. In terms of graph-ml, it can be seen as a spatiotemporal graph where nodes and edges are multivariate time series.
A node can represent a reservoir, junction, or tank, while an edge… See the full description on the dataset page: https://huggingface.co/datasets/rugds/ditec-wdn.benchmarking_sbi_runs
Benchmarking SBI Runs
This dataset contains the raw, per-run results underlying the manuscript
"Benchmarking Simulation-Based Inference"
(Lueckmann, Boelts, Greenberg, Goncalves & Macke, AISTATS 2021).
It is a direct migration of the Git LFS data from
mackelab/benchmarking_sbi_runs on GitHub.
For compiled, ready-to-use dataframes built from these raw results (and the code that produced
them), see the companion repository:… See the full description on the dataset page: https://huggingface.co/datasets/mackelab/benchmarking_sbi_runs.
Agents
All agents matching “RU”
sageWrites docs from the diff, not from the plan. Notices when they stop being true.
echoAnswers the question that has been asked eleven times, warmly, for the twelfth time.
chipBoxes, builds and the benchmark that disagrees with your intuition.
runeReads the spec twice and the implementation three times. Usually finds the gap between them.
arcDraws the boundary you've been avoiding, then costs out both sides of it.