datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
imagenet_hard_review_data_r2Ko-Agent-Trajectories-1.0
Ko-Agent-Trajectories-1.0
Dataset card v1.1.1 (2026-09-22). The pipeline code is now released in this repository
under pipeline/, together with the API catalogue, the scenario templates and the complete
prompt set. The card reports the completed human review study and the v1.1 artefacts
(behaviour DPO config, per-item validation scores, manifest, filter asset).
Korean edition: README.ko.md.
TL;DR
A Korean multi-turn agent ↔ tool trajectory corpus synthesized… See the full description on the dataset page: https://huggingface.co/datasets/taejoon89/Ko-Agent-Trajectories-1.0.hanabi-zscfinal-qwen3-8b-logs
zsc_final Qwen3-8B 실행 로그 — 서버 10.45.7.134
llm_agent 저장소 zsc_final 에서 2026-09-21~22 에 이 서버가 돌린 것들의
로그다. 수치의 정본은 저장소의 zsc_final/result_loop_trial.md 와
zsc_final/result_hidden_store.md 다.
무엇이 들어 있나
폴더
무엇
logs/ev_*.log
LLM 게임 세 조건 (A 셀 · scope 셀 · C 셀), 판 3000~3059, 조건당 LLM 게임 240개
logs/store_game_s*.log
은닉 저장소 1차 실행 (조각 셋)
logs/store_game_t*.log
은닉 저장소 본 실행 (조각 여섯), LLM 게임 240개
logs/hdr_cut2.log
헤더 토막을 빼며 판독 정확도를 잰 실행
store_index/index_*of6.jsonl… See the full description on the dataset page: https://huggingface.co/datasets/taekbae/hanabi-zscfinal-qwen3-8b-logs.hanabi-zscfinal-hidden-store-qwen3-8b
Hanabi run_v1 LLM 게임 은닉 저장소 — Qwen3-8B, A 셀(raw)
llm_agent 저장소 zsc_final 의 src/env2_hidden_store.py (git 107a5a8)
로 2026-09-22 에 서버 10.45.7.134 에서 뽑았다. 수치의 정본은 저장소의
zsc_final/result_hidden_store.md 다.
무대
Qwen3-8B. 측정 틀 run_v1 — p_v2 프롬프트 헤더, 발신은 양 자리 모두 대본
(규칙대로만 낸다), 수신만 LLM. 조건은 A 셀 (raw, 수신에 아무 문장도 안
넣는 무개입). 판 3000~3059 가 LLM 게임 판이고, 판 하나가 규약 구성
4개에서 돌아 LLM 게임 4개가 된다. 판 60 × 구성 4 = LLM 게임 240.
파일
raw_c{구성}_e{판}_s{좌석}.npz — LLM 게임 하나에 좌석 둘, 합쳐… See the full description on the dataset page: https://huggingface.co/datasets/taekbae/hanabi-zscfinal-hidden-store-qwen3-8b.imagenet_hard_review_data
