datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llm-serving-selector-regret
LLM-Serving Selector Regret
LLM-Serving Selector Regret is a metrics-only research dataset for studying learned policy selection in LLM-serving schedulers. It contains derived selector/oracle/regret objects generated by Soroush Vahidi's research workflow, not raw request traces.
Creator / Provider
Dataset creator/provider: Soroush Vahidi.
The released selector/regret and policy-suitability metrics were generated by Soroush Vahidi's research workflow. Underlying… See the full description on the dataset page: https://huggingface.co/datasets/SoroushVahidi/llm-serving-selector-regret.repro-on-regret-bounds-of-thompson-sampling-for-bayesian-optimization-traces
Agent traces
Agent sessions published from a Trackio Logbook.
repro-optimal-regret-for-policy-optimization-in-contextual-bandits-traces
Agent traces
Agent sessions published from a Trackio Logbook.
repro-dynamic-regret-via-discounted-to-dynamic-reduction-with-applications-to-curved-l-traces
Agent traces
Agent sessions published from a Trackio Logbook.
repro-finite-and-corruption-robust-regret-bounds-in-online-inverse-linear-optimization-traces
Agent traces
Agent sessions published from a Trackio Logbook.
repro-tighter-regret-lower-bound-for-gaussian-process-bandits-with-squared-exponential-traces
Agent traces
Agent sessions published from a Trackio Logbook.
