llm-trajectories
llm-eval-flakiness-trajectories
Llm Eval Flakiness Trajectories
Rights & intended use: legacy public research corpus / portfolio
artifact. Hosted frontier-model outputs are research-only inputs under
project policy (synthetic-factory#161):
intended_use: research_only, project_training_policy: blocked. Not
training data for any model-weight update. Machine-readable record:
rights.json.
Release status: The raw, uncurated payload is now published under
data/raw/. It is available for inspection and… See the full description on the dataset page: https://huggingface.co/datasets/rmems/llm-eval-flakiness-trajectories.evo_llm_trajectories
What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search
This dataset contains optimization trajectories for 15 Large Language Models (LLMs) across 8 different optimization tasks, as presented in the paper What Makes an LLM a Good Optimizer? A Trajectory Analysis of LLM-Guided Evolutionary Search.
The data was collected using the LLMEvo framework to study how various LLMs behave when orchestrating evolutionary and agentic optimization systems. The… See the full description on the dataset page: https://huggingface.co/datasets/LivevreXH/evo_llm_trajectories.llm-deception-trajectories
LLM Deception Trajectories
Hidden-state trajectories from 11 transformer architectures processing matched truthful/deceptive prompt pairs across 20 deception categories.
Dataset Description
This dataset captures the internal processing trajectories of large language models as they generate responses to truthful vs. deceptive prompts. Each trajectory records the hidden state at every transformer layer, enabling analysis of how deception manifests in model… See the full description on the dataset page: https://huggingface.co/datasets/dSLLab/llm-deception-trajectories.llm-srbench-trajectories
