lihaoxin2020/evidence-subagent-sft-gpt54-single-all-jina-v2-qwen35-thinking
Evidence Subagent SFT GPT-5.4 Jina v2, Qwen3.5 Thinking Aligned This dataset is an aligned version of lihaoxin2020/evidence-subagent-sft-gpt54-single-all-jina-v2 for supervised fine-tuning a Qwen3.5 evidence subagent in LLaMA-Factory. Splits train: 10,379 examples validation: 100 examples Format Each row contains: id: source trajectory id conversations: OpenAI-style messages with roles system, user, function, tool, and assistant tools:… See the full description on the dataset page: https://huggingface.co/datasets/lihaoxin2020/evidence-subagent-sft-gpt54-single-all-jina-v2-qwen35-thinking.
Evidence Subagent SFT GPT-5.4 Jina v2, Qwen3.5 Thinking Aligned
This dataset is an aligned version of lihaoxin2020/evidence-subagent-sft-gpt54-single-all-jina-v2 for supervised fine-tuning a Qwen3.5 evidence subagent in LLaMA-Factory.
Splits
train: 10,379 examplesvalidation: 100 examples
Format
Each row contains:
id: source trajectory idconversations: OpenAI-style messages with rolessystem,user,function,tool, andassistanttools: JSON-encoded OpenAI-compatible tool schemasmetadata: source and trajectory metadata
The trajectory shape is:
system -> user -> function -> tool -> ... -> assistantThe function and final assistant messages preserve teacher thinking blocks normalized for the LLaMA-Factory qwen3_5 template:
<think>
...
</think>
...The final assistant message contains exactly one <evidence_report>...</evidence_report> block.
LLaMA-Factory
Use the dataset as an OpenAI-format dataset with the qwen3_5 template. The local generation command used to prepare this upload was:
python agent/scripts/prepare_evidence_subagent_llamafactory.py --keep-think