CoolFace
20 results

albert

albertvillanova /datasets-tests-compressiontextn<1K0 likes64k downloads5y agoHugging Facealbertvillanova /tests-raw-jsonltext10K<n<100K1 likes40k downloads5y agoHugging Facealbertklorer /safedocs-1M-muse-spark-1.3-judged SafeDocs: Muse Spark 1.3 judge annotations Incrementally published, one complete shard per commit. All original source columns, images, complete Paddle JSON, rows and row order are preserved. No language or quality filtering. New columns: judge_verdict (PERFECT/ERROR), judge_reason, judge_status, and judge_error. Operational failures retain the original page with a null verdict and reason, status failed, and a diagnostic in judge_error; they are not OCR ERRORs. Direct Meta API… See the full description on the dataset page: https://huggingface.co/datasets/albertklorer/safedocs-1M-muse-spark-1.3-judged.tabular100K<n<1M0 likes12k downloads2d agoHugging Facealbertwilcox /mesa-all-train-lerobot0 likes4.2k downloads6mo agoHugging Facealbertklorer /safedocs-1M1 likes3.6k downloads1d agoHugging FacealbertoRodriguez97 /history-anchor-100-traces History Anchor 100 — Model Trajectories *Per-(model × condition × scenario set × seed) raw outputs from the paper "History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions".* This dataset contains the full set of model decisions that back every figure and table in the paper. Use it to: audit a single model's behaviour scenario-by-scenario, recompute headline metrics without re-running the (paid) API sweeps, mine reasoning_content traces from models that expose… See the full description on the dataset page: https://huggingface.co/datasets/albertoRodriguez97/history-anchor-100-traces.text-generation10K<n<100K0 likes1.7k downloads4mo agoHugging Face