JetBrains-Research/lca-codegen-large
LCA Project Level Code Completion How to load the dataset from datasets import load_dataset ds = load_dataset('JetBrains-Research/lca-codegen-large', split='test') Data Point Structure repo – repository name in format {GitHub_user_name}__{repository_name} commit_hash – commit hash completion_file – dictionary with the completion file content in the following format: filename – filepath to the completion file content – content of the completion… See the full description on the dataset page: https://huggingface.co/datasets/JetBrains-Research/lca-codegen-large.
032
../
test-00000-of-00006-2bf99ae1b9916fc8.parquetdownload
test-00001-of-00006-8130758795d39c4e.parquetdownload
test-00002-of-00006-f2ff77a6b58ff3f6.parquetdownload
test-00003-of-00006-309977fe7911f160.parquetdownload
test-00004-of-00006-ceea4952757bdcc8.parquetdownload
test-00005-of-00006-e6e02a71c4a4d6c4.parquetdownload
