Team Ai
Datasetpublic

AISE-TUDelft/multilingual-code-comments-fixed-8

Multilingual code comments This dataset contains 500 source-code examples for each of Chinese, Dutch, English, Greek and Polish (2,500 examples total), with comments generated by five models and human correctness ratings and error annotations. Each language has a train split. Human ratings use Correct, Partial and Incorrect. Error fields contain comma-separated taxonomy codes. A generated comment can carry multiple error codes. Annotation interpretation Stored… See the full description on the dataset page: https://huggingface.co/datasets/AISE-TUDelft/multilingual-code-comments-fixed-8.

sourceHugging Faceupdated 7h agoView on Hugging Face
0likes163downloads
../
filetrain-00000-of-00001.parquet8.6 MBdownload

AISE-TUDelft/multilingual-code-comments-fixed-8 · main · files are served by the source, never re-hosted here