nl2bash
Datasets
All datasets matching “nl2bash”nl2bash-custom
nl2bash-custom
nl2bash-custom is a custom dataset used to fine-tune Large Language Models for Bash Code Generation. Fine tune the Code-Llamma family of LLMs (7b, 13b, 70b) for best results.
The dataset is created by reformatting and reshiffling of 2 original datasets
nl2bash by TelinaTool
NLC2CMD by Magnum Reasearch Group
Dataset Structure
train.json: Training split.
dev.json: Development split.
test.json: Test split.
Usage
from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/AnishJoshi/nl2bash-custom.rl__24GPU_base__mix_h2_language_balanced__r2egym-nl2bash-stackrl__24GPU_base__swe_rebench_patched_oracle__r2egym-nl2bash-stackNL2Bash
NL2Bash
This dataset is a collection of natural language (English) instructions and corresponding Bash commands for the task of natural language to Bash translation (NL2Bash).
Dataset Details
Dataset Description
Our dataset is enhanced version of NL2SH-ALFA dataset. We improved the natural language instruction to make them more clear to avoid ambuiguity. NL2SH-ALFA dataset was produced by combining, deduplicating and filtering multiple datasets from… See the full description on the dataset page: https://huggingface.co/datasets/dilkushsingh/NL2Bash.nl2bashThe dataset is constructed from
https://github.com/TellinaTool/nl2bashDCAgent2_terminal_bench_2_DCAgent2_nl2bash-verified-GLM-4.6-traces-32ep-32k_glof5f5a963
