Team Ai
Datasetpublic

google-research-datasets/great_code

The dataset for the variable-misuse task, described in the ICLR 2020 paper 'Global Relational Models of Source Code' [https://openreview.net/forum?id=B1lnbRNtwr] This is the public version of the dataset used in that paper. The original, used to produce the graphs in the paper, could not be open-sourced due to licensing issues. See the public associated code repository [https://github.com/VHellendoorn/ICLR20-Great] for results produced from this dataset. This dataset was generated synthetically from the corpus of Python code in the ETH Py150 Open dataset [https://github.com/google-research-datasets/eth_py150_open].

sourceHugging Facecc-by-sa-3.0updated 3y agoView on Hugging Face
5likes413downloads
settings

This repository belongs to google-research-datasets on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namegreat_code
visibilitypublic
licencecc-by-sa-3.0
gatedno
ownergoogle-research-datasets
Account settings
google-research-datasets/great_code · Team Ai