Team Ai
Datasetpublic

CreativeLang/ColBERT_Humor_Detection

ColBERT_Humor Dataset Summary ColBERT Humor contains 200,000 labeled short texts, equally distributed between humorous and non-humorous content. The dataset was created to overcome the limitations of prior humor detection datasets, which were characterized by inconsistencies in text length, word count, and formality, making them easy to predict with simple models without truly understanding the nuances of humor. The two sources for this dataset are the News… See the full description on the dataset page: https://huggingface.co/datasets/CreativeLang/ColBERT_Humor_Detection.

sourceHugging Facecc-by-2.0updated 3y agoView on Hugging Face
7likes79downloads
dataset.csv4 linesDownload Raw Back to root
1version https://git-lfs.github.com/spec/v12oid sha256:5b48647ebdf44a54f0a29c1011c5d397b013057111ed215a8382616f513ff8843size 148742404