Team Ai
Datasetpublic

EngineeringSoftware/PLSemanticsBench

The 43rd International Conference on Machine Learning (ICML 2026), Seoul, South Korea LLMs Lean on Priors, Not Programming Language Semantics by Aditya Thimmaiah1, Jiyang Zhang1, Jayanth Srinivasa2, Junyi Jessy Li1, Milos Gligoric1 1The University of Texas at Austin     2Cisco Research TLDR: Frontier LLMs execute programs with up to 90–100% accuracy when symbols retain their usual… See the full description on the dataset page: https://huggingface.co/datasets/EngineeringSoftware/PLSemanticsBench.

sourceHugging Facecc-by-4.0updated 6d agoView on Hugging Face
1likes152downloads
../
fileK_NonStandard_NumDescription5-00000-of-00001.parquet36 KBdownload
fileK_Standard_NumDescription5-00000-of-00001.parquet32 KBdownload
fileS_NonStandard_NumDescription5-00000-of-00001.parquet42 KBdownload
fileS_Standard_NumDescription5-00000-of-00001.parquet33 KBdownload

EngineeringSoftware/PLSemanticsBench · main · files are served by the source, never re-hosted here