InfoBayAI/Indonesian-STEM-Textbook-Dataset
Dataset Description: This dataset is a large-scale collection of Indonesian STEM textbook data, containing 5,169 books and 208.30 million words, designed to support the development and training of advanced NLP systems and AI models for scientific understanding, problem-solving, and concept learning in Bahasa. Full Dataset Overview This dataset is part of a large-scale multilingual educational corpus containing over 3+ billion words across 5,000+ subjects, supported by interwoven images for… See the full description on the dataset page: https://huggingface.co/datasets/InfoBayAI/Indonesian-STEM-Textbook-Dataset.
035
No card is published for this repository, or it could not be fetched from Hugging Face right now.
