Team Ai
Datasetpublic

ajibawa-2023/Shell-Code-Large

Shell-Code-Large Shell-Code-Large is a large-scale corpus of Shell scripting source code comprising approximately 640,000 code samples stored in JSON Lines (.jsonl) format. The dataset is designed to support research in large language model (LLM) pretraining, code intelligence, DevOps automation, cloud infrastructure engineering, system administration, and software engineering automation. By providing a high-volume, language-specific corpus focused exclusively on Shell scripting… See the full description on the dataset page: https://huggingface.co/datasets/ajibawa-2023/Shell-Code-Large.

sourceHugging Facemitupdated 4mo agoView on Hugging Face
21likes132downloads
5 commits on main
91adad64mo ago

Update README.md

ajibawa-2023
25e7c754mo ago

Update README.md

ajibawa-2023
525cecc4mo ago

Update README.md

ajibawa-2023
8ad54b74mo ago

Upload 6 files

ajibawa-2023
789cc324mo ago

initial commit

ajibawa-2023