Team Ai
Datasetpublic

EleutherAI/wikitext_document_level

Wikitext Document Level This is a modified version of https://huggingface.co/datasets/wikitext that returns Wiki pages instead of Wiki text line-by-line. The original readme is contained below. Dataset Card for "wikitext" Dataset Summary The WikiText language modeling dataset is a collection of over 100 million tokens extracted from the set of verified Good and Featured articles on Wikipedia. The dataset is available under the Creative Commons… See the full description on the dataset page: https://huggingface.co/datasets/EleutherAI/wikitext_document_level.

sourceHugging Facecc-by-sa-3.0updated 2y agoView on Hugging Face
18likes87kdownloads

Nothing at this path on main. The folder may be empty, or the revision may not exist.

EleutherAI/wikitext_document_level · main · files are served by the source, never re-hosted here