Team Ai
Datasetpublicgated

InfoBayAI/Arabic-STEM-Textbook-Dataset

Dataset Description: This dataset is a large-scale collection of Arabic STEM textbook data, containing 1,364 books and 63.51 million words, designed to support the development and training of advanced NLP systems and AI models for scientific understanding, problem-solving, and concept learning in Arabic. Full Dataset Overview This dataset is part of a large-scale multilingual educational corpus containing over 3+ billion words across 5,000+ subjects, supported by interwoven images for deeper… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Arabic-STEM-Textbook-Dataset.

sourceHugging Facecc-by-4.0updated 2d agoView on Hugging Face
0likes24downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.