CoolFace
17 results

self-evolving

thaki-AI /daily-paper-2026-08-29-goodhart-shift-self-evolving-harness The Goodhart Shift: Measuring Train-vs-Holdout Divergence in Overnight Self-Evolving Agent Skill Loops TL;DR — When an overnight self-evolving agent loop edits its own skills against the same bench it reports on, its self-reported gains can stop generalizing. This paper formalizes that failure as the Goodhart shift - a per-night train-vs-holdout divergence - and specifies a sealed-holdout divergence gate that tells "the loop improved" apart from "the loop overfit." ThakiCloud… See the full description on the dataset page: https://huggingface.co/datasets/thaki-AI/daily-paper-2026-08-29-goodhart-shift-self-evolving-harness.0 likes161 downloads25d agoHugging FaceBin-Wu /self-evolving-simulationtext10M<n<100M0 likes83 downloads3mo agoHugging Facexunyoyo /Self-Evolving-Safety The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies Paper | Project Page This repository contains the dataset and empirical results for the paper "The Devil Behind Moltbook: Anthropic Safety is Always Vanishing in Self-Evolving AI Societies". Summary This research investigates the self-evolution trilemma in multi-agent systems built from large language models (LLMs). The authors demonstrate both theoretically and empirically that… See the full description on the dataset page: https://huggingface.co/datasets/xunyoyo/Self-Evolving-Safety.text-classification0 likes53 downloads7mo agoHugging Face11-47 /self_evolving_self_debugging_250_implementations-2textn<1K1 likes24 downloads9mo agoHugging Facehyojuuun /self_evolving_iter-qwen-qwen3-4b-base_math_1223_0510-v1text1K<n<10K0 likes22 downloads9mo agoHugging Facehyojuuun /self_evolving_iter-qwen-qwen3-4b-base_olympiad_1226_0513-v3text1K<n<10K0 likes22 downloads9mo agoHugging Face