Team Ai
Datasetpublic

AlgorithmicResearchGroup/arxiv_deep_learning_python_research_code

ArXiv Deep Learning Python Research Code A curated corpus of Python source code files extracted from GitHub repositories referenced in ArXiv papers. Contains 391,496 files (1.49 GB) filtered to deep learning frameworks, designed for training and evaluating Code LLMs on research-grade code. Dataset Summary Statistic Value Total files 391,496 Total size 1.49 GB Source repos 34,099 Time span ArXiv inception through July 2023 Dataset… See the full description on the dataset page: https://huggingface.co/datasets/AlgorithmicResearchGroup/arxiv_deep_learning_python_research_code.

sourceHugging Faceotherupdated 6mo agoView on Hugging Face
10likes168downloads
settings

This repository belongs to AlgorithmicResearchGroup on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namearxiv_deep_learning_python_research_code
visibilitypublic
licenceother
gatedno
ownerAlgorithmicResearchGroup
Account settings
AlgorithmicResearchGroup/arxiv_deep_learning_python_research_code · Team Ai