Team Ai
Datasetpublic

MehdiAstaraki/progressive-code-switch

MehdiAstaraki/progressive-code-switch Progressive code-switching retrieval-decay benchmark. Each base patent yields a cumulative ladder of documents (base__r0 clean → base__rN), where each step swaps one more chemistry term into another language / spelling / ChEBI form. One fixed question per base (about the step-1 term) is reused for every depth; the qrels carry the depth so you can measure how retrieval decays as more terms are code-switched. Configs: corpus (ladder variant… See the full description on the dataset page: https://huggingface.co/datasets/MehdiAstaraki/progressive-code-switch.

sourceHugging Faceupdated 4mo agoView on Hugging Face
0likes4downloads
5 commits on main
0a3fffe4mo ago

Upload README.md with huggingface_hub

MehdiAstaraki
0348d144mo ago

Upload dataset

MehdiAstaraki
51fe5214mo ago

Upload dataset

MehdiAstaraki
1dcc16d4mo ago

Upload dataset

MehdiAstaraki
6a5f1784mo ago

initial commit

MehdiAstaraki