datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fineproofs-prm-context-v2-full-cot-xprob
FineProofs PRM Context v2: Full Cot Xprob
This arm uses full_cot context from cross_problem rollouts with packing policy middle_truncated_reasoning_equal_share.
This is one of nine row-matched context variants built from the verified FineProofs rollout collection. Partial-prefix and complete-response targets both use canonical normalized rubric credit derived from clamped points divided by max points. The correct column is only a legacy boolean projection at reward >= 0.5;… See the full description on the dataset page: https://huggingface.co/datasets/asingh15/fineproofs-prm-context-v2-full-cot-xprob.prm-mc-value-context-full-cot
MC-value PRM dataset — context mode: full_cot
Data for training an in-context / policy-conditioned Monte-Carlo value PRM. Each row's query is a
partial reasoning prefix; the target reward = V = P(correct | prefix), the Monte-Carlo value estimated
from branched Qwen3.5-4B rollouts on Polaris math problems. The user prompt additionally carries a
"# Other attempts by the same model at this problem" block — the ablation variable.
Context for this variant: Up to K=4 OTHER attempts'… See the full description on the dataset page: https://huggingface.co/datasets/asingh15/prm-mc-value-context-full-cot.fineproofs-prm-context-v2-full-cot
FineProofs PRM Context v2: Full Cot
This arm uses full_cot context from same_problem rollouts with packing policy middle_truncated_reasoning_equal_share.
This is one of nine row-matched context variants built from the verified FineProofs rollout collection. Partial-prefix and complete-response targets both use canonical normalized rubric credit derived from clamped points divided by max points. The correct column is only a legacy boolean projection at reward >= 0.5; training… See the full description on the dataset page: https://huggingface.co/datasets/asingh15/fineproofs-prm-context-v2-full-cot.superskillret-index-fullcontext
superskillret prebuilt index — full-context
Prebuilt embedding index for the superskillret Claude Code plugin.
Unlike the default index (which embeds only name + description), this build
encodes the full skill body (name + description + body) up to
max_seq_length=32768 tokens. Larger index, much higher recall on
skills whose name/description don't capture every keyword in the body.
Version: 1
Corpus: ThakiCloud/SKILLRET (train+test)
Encoder: ThakiCloud/SkillRet-Embedding-0.6B… See the full description on the dataset page: https://huggingface.co/datasets/youngryankim/superskillret-index-fullcontext.full-single-contextfull-pipeline-context-results-allcontextual_code_review_fullmem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c8192-t4096-1000s-a-fullcontextsql-create-context-full
Dataset Card for "sql-create-context-full"
More Information needed
mem_agent-model_based-rl-memoryagent-7b-barexamqa-train-c256-t128-1000s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-triviaqa-llama-memorization-val-c4096-t2048-fullcontextmem_agent-model_based-rl-memoryagent-7b-triviaqa-val-c4096-t2048-1000s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-docfinqa-train-c4096-t4096-1000s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-barexamqa-train-c256-t128-20s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-infbench-longbook-qa-test-c8192-t4096-20s-ag-fullcontextmem_agent-model_based-rl-memoryagent-7b-infbench-longbook-qa-test-c8192-t4096-10s-ag-fullcontextmem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c8192-t4096-10s-agn-fullcontextmem_agent-model_based-rl-memoryagent-7b-ruler-qa-test-c2048-t1024-20s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-housingqa-test-c1024-t512-1000s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-infbench-longbook-qa-test-c8192-t4096-1000s-fullcontextmem_agent-model_based-rl-memoryagent-7b-housingqa-test-c1024-t512-20s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c8192-t4096-20s-agn-fullcontextmem_agent-model_based-rl-memoryagent-7b-docfinqa-train-c4096-t4096-20s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-infbench-code-debug-test-c4096-t4096-20s-agn-fullcontextmem_agent-model_based-rl-memoryagent-7b-infbench-longbook-qa-test-c4096-t4096-20s-ag-fullcontextmem_agent-model_based-rl-memoryagent-7b-ruler-qa-test-c2048-t1024-10s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-barexamqa-train-c256-t128-10s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-gsminf-ops_12-c4096-t2048-20s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-bizbench-test-c1024-t512-1000s-agnostic-fullcontextmem_agent-model_based-rl-memoryagent-7b-ruler-qa-test-c2048-t1024-1000s-agnostic-fullcontext
