strategy-scope/si_et_al-ideation-gpt41mini-20260414_233911-ideas
si_et_al-ideation-gpt41mini-20260414_233911-ideas Per-idea flat table with LLM-judge scores. Parent: si_et_al-ideation-gpt41mini-20260414_233911 Columns topic: NLP topic idea_index: position within pooled topic idea set run_idx: which of the n_runs generation calls produced this idea idea_text: the generated idea (Problem/Existing Methods/Motivation/Proposed Method/Experiment Plan) overall: 1-10 LLM-judge score (port of ai_researcher/src/idea_direct_score.py)… See the full description on the dataset page: https://huggingface.co/datasets/strategy-scope/si_et_al-ideation-gpt41mini-20260414_233911-ideas.
03
sietal-ideation-gpt41mini-20260414_233911-ideas
Per-idea flat table with LLM-judge scores. Parent: si_et_al-ideation-gpt41mini-20260414_233911
Columns
- topic: NLP topic
- idea_index: position within pooled topic idea set
- run_idx: which of the n_runs generation calls produced this idea
- idea_text: the generated idea (Problem/Existing Methods/Motivation/Proposed Method/Experiment Plan)
- overall: 1-10 LLM-judge score (port of airesearcher/src/ideadirect_score.py)
- tournament_score: Swiss tournament pairwise score (port of tournament_ranking.py); only present if --tournament was used
- raw_judge_response: raw judge output (useful for debugging parse failures)
