marin-community/open-thoughts-4-30k-code-qwen3-32b-annotated-32768-tokens
Dataset Card for Open-Thoughts-4-30K-Code-Qwen3-32B-Annotated-32768-Tokens Overview This dataset is a variant of marin-community/open-thoughts-4-30k-code-qwen3-32b-annotated with an extended maximum sequence length. The responses in the generated_text column were generated with max output tokens = 32768 (instead of 7500 in the original dataset), allowing for longer and more complete chain-of-thought reasoning. Generation Details Model:… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-code-qwen3-32b-annotated-32768-tokens.
Add dataset card with metadata
Add dataset card
Transform dataset: remove response_seed and messages columns, add conversations column
Uploading small files from gs://marin-us-central1/documents/open-thoughts-4-30k-code-qwen3-32b-annotated-o32768-8b7173
Commiting these files to the repo from gs://marin-us-central1/documents/open-thoughts-4-30k-code-qwen3-32b-annotated-o32768-8b7173:
Commiting these files to the repo from gs://marin-us-central1/documents/open-thoughts-4-30k-code-qwen3-32b-annotated-o32768-8b7173:
initial commit
