MoreThought/DeepSWEGym
Dataset Description This dataset is a filtered version of all the SWE-bench/SWE-smith-lang datasets (expect php) merged together (originally 88k rows total), it aims to improve benchmark results on DeepSWE-style problems, benchmarks, and general coding skills. It is specifically filtered for rows with complex/long code problems in the original datasets, having an average row size of 94.41kb, a total uncompressed size of 6.12GB, and a total of 64821 examples. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/MoreThought/DeepSWEGym.
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Rename README (1).md to README.md
Upload README (1).md
Delete README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Create README.md
Rename filtered_swe_smith.jsonl to dataset.jsonl
Upload filtered_swe_smith.jsonl with huggingface_hub
initial commit
