Team Ai
Datasetpublic

ccdv/arxiv-classification

Arxiv Classification: a classification of Arxiv Papers (11 classes). This dataset is intended for long context classification (documents have all > 4k tokens). Copied from "Long Document Classification From Local Word Glimpses via Recurrent Attention Learning" @ARTICLE{8675939, author={He, Jun and Wang, Liqun and Liu, Liu and Feng, Jiao and Wu, Hao}, journal={IEEE Access}, title={Long Document Classification From Local Word Glimpses via Recurrent Attention Learning}, year={2019}… See the full description on the dataset page: https://huggingface.co/datasets/ccdv/arxiv-classification.

sourceHugging Faceupdated 2y agoView on Hugging Face
27likes2.3kdownloads
17 commits on main
d4bf86d2y ago

Convert dataset to Parquet (#4)

albertvillanova
f9bd9214y ago

Fix language and task array (#2)

ccdv, albertvillanova
6ec9f334y ago

for windows

ccdv
6d2e2a24y ago

readme

ccdv
ba431444y ago

update

ccdv
63ace995y ago

Update README.md

ccdv
4820c8f5y ago

fix

tkon3
a5930d95y ago

test

tkon3
41ae51d5y ago

fix

tkon3
e9da07e5y ago

refix

tkon3
e09f2775y ago

fix

tkon3
e41a1f55y ago

update

tkon3
e5a36965y ago

fix

tkon3
168b6555y ago

fix

tkon3
31eaed95y ago

fix

tkon3
3d71fdf5y ago

first commit

tkon3
2ac48e85y ago

initial commit

system