G-reen/encoder-decoder-trial-stat
Encoder/decoder trial: encoder-marginal report Dataset: G-reen/encoder-decoder-trial-stat Rows analysed: 122,933 (every kept (encoder, decoder, source row) triple; source G-reen/cc-re-2021-filtered shard 0, 2000 rows of at most 4000 words) Prompt file: prompts/indirect_reference_dataset_train.json (turn 0 encodes the document, turn 1 reconstructs it from the encoding alone) Encoders: 9 (granite-4.2-30b-nvfp4 [0], Ornith-1.5-35B-A3B-NVFP4 [1], Llama-3.3-70B-Instruct-NVFP4 [2]… See the full description on the dataset page: https://huggingface.co/datasets/G-reen/encoder-decoder-trial-stat.
0703
1version https://git-lfs.github.com/spec/v12oid sha256:b7fcf36ffa0e2da1fa8e2036273377c7c2ed78bcd551cf66d7ed7e290d3a48d53size 1601894 