robustness
llama-3-8b-Instruct-bnb-4bit-robustness_primaryllama-3-8b-Instruct-bnb-4bit-robustness_latestrt-sam.backdoor_9_lr6e-5_rho0.1z0406_bt_ordinary_RT_1e-4_bt_backdoor_0_lr3e-6_piratez0406_bt_ordinary_RT_3e-4_bt_backdoor_0_lr3e-6_alpacarobustness_t5rt-sam.backdoor_9_lr1e-5_rho0.01z0406_rt_broad_RT_backdoor_0_lr1e-6
kitti-yolo11n-robustness-benchmark
KITTI YOLO11n Robustness & Adversarial Benchmark Suite
This dataset contains 649,425 benchmark samples evaluating the perception robustness of YOLO11n (Ultralytics YOLOv11 nano in original FP32 precision) on the official KITTI Object Detection train set (3,711 images) under 35 attack & corruption techniques across 5 severity levels.
?? Benchmark Leaderboard (mAP@0.5 Drop on YOLO11n)
Clean Baseline AP50: 0.3555
Evaluation Model: YOLO11n (Original weights:… See the full description on the dataset page: https://huggingface.co/datasets/VietPhong/kitti-yolo11n-robustness-benchmark.staining-robustness-evaluation
A Protocol for Evaluating Robustness to H&E Staining Variation in Computational Pathology Models
This repository provides the stain references, pretrained models, and experimental results required to:
Define custom staining references using our PLISM reference library
Reproduce our published controlled staining robustness experiments
👉 Code repository: https://github.com/lely475/staining-robustness-evaluation/tree/main
👉 Associated publication: Paper
Overview: How… See the full description on the dataset page: https://huggingface.co/datasets/CTPLab-DBE-UniBas/staining-robustness-evaluation.lighting-invariant-bedroom-perception-robustness-benchmark
Lighting-Invariant Bedroom Perception & Robustness Benchmark
Generated by datapack-import.ts
This dataset mirrors public data-pack render outputs from Physicl.
Each row represents one render view. The image column contains a stable URL to the primary render image uploaded under /data; image_path stores the relative repository path and data_commit_sha pins the Hugging Face dataset commit used by those URLs. Files are uploaded as downloaded unless optional PNG recompression is… See the full description on the dataset page: https://huggingface.co/datasets/physicl/lighting-invariant-bedroom-perception-robustness-benchmark.tokamark-robustness-data
TokaMark Sensor Robustness Benchmark Data
Associated paper: Benchmarking Sensor Robustness in Plasma Diagnostic Models: A Systematic Evaluation on TokaMarkAuthor: Neerav GuptaCode: github.com/Neerav-Gupta/tokamark-robustness
Dataset Description
This dataset contains pre-processed numpy arrays, trained model checkpoints, and experiment results from the first systematic robustness benchmark of plasma diagnostic ML models under realistic sensor failure, using the… See the full description on the dataset page: https://huggingface.co/datasets/Neerav-Gupta/tokamark-robustness-data.bangla-noise-robustness-dataeval_robustness_e9_3_full_dp_local_180k_fThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "koch",
"total_episodes": 175,
"total_frames": 73950,
"total_tasks": 1,
"total_videos": 350,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:175"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/nduque/eval_robustness_e9_3_full_dp_local_180k_f.
