Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01mib-bench /copycolors_mcqaThis dataset consists of formatted n-way multiple choice questions, where n is in [2,10]. The task itself is simply to copy the prototypical color from the context and produce the corresponding color's answer choice letter. The "prototypical colors" dataset instances themselves come from Memory Colors (Norland et al. 2021) and corypaik/coda (instances whose object_group is 0, indicating participants agreed on a prototypical color of that object). tabularquestion-answering1K<n<10K0 likes2.2k downloads2y agoHugging Face02nvidia /Nemotron-RL-knowledge-mcqa Dataset Description: The Nemotron-RL-knowledge-mcqa is a multi-domain synthetic multiple-choice question-answering (MCQA) dataset containing knowledge based questions. It combines and refines subsets of the [OpenScienceReasoning-2] (https://huggingface.co/datasets/nvidia/OpenScienceReasoning-2) dataset and other unstructured sources such as books and articles.The dataset was created using Qwen3-32B, [Qwen3-235B-A22B-Instruct-2507]… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-knowledge-mcqa.text100K<n<1M13 likes2.1k downloads12d agoHugging Face03lighteval /med_mcqaFrom "MedMCQA: A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering" (Pal et al.), MedMCQA is a "multiple-choice question answering (MCQA) dataset designed to address real-world medical entrance exam questions." The dataset "...has more than 194k high-quality AIIMS & NEET PG entrance exam MCQs covering 2.4k healthcare topics and 21 medical subjects are collected with an average token length of 12.77 and high topical diversity." The following is an example from… See the full description on the dataset page: https://huggingface.co/datasets/lighteval/med_mcqa.text100K<n<1M13 likes864 downloads3y agoHugging Face04ESmike /driving_mcqa DrivingExamMCQA The DrivingExamMCQA dataset is a Multiple-Choice Question Answering (MCQA) collection based on real driving exam questions. It supports multilingual assessment across three languages: Arabic (ar), French (fr), and English (en) (with translations). Overview Each language includes two modalities: Image-supported questions (_img splits): Questions paired with an image (e.g., road signs, traffic scenarios). Text-only questions (_text splits): Standard… See the full description on the dataset page: https://huggingface.co/datasets/ESmike/driving_mcqa.imagemultiple-choicen<1K2 likes784 downloads11mo agoHugging Face05Emulated-Inc /science-mcqa-training-pool Science multiple-choice training pool Public multiple-choice science questions from three datasets, read at the pinned revisions named below and laid out twice. Train on either layer or on both. pool.jsonl Every source rewritten into one shape, 182035 rows, one JSON object per line, with these fields. Field What it holds id a row identifier unique within this file question the question text, as its source publishes it options the answer options, as… See the full description on the dataset page: https://huggingface.co/datasets/Emulated-Inc/science-mcqa-training-pool.textquestion-answering100K<n<1M0 likes704 downloads17d agoHugging Face06aisingapore /NLU-Belebele-MCQAgatedtext10K<n<100K0 likes641 downloads10mo agoHugging Face07andresnowak /MNLP_MCQA_datasetThis MCQA dataset (of only single answer) contains a mixture of train, validation and test from this datasets (test and validation are only used for testing not for training): mmlu auxiliary train Only the stem subset is used mmlu Only the stem subset is used mmlu 10 choices auxiliary train stem ai2_arc ScienceQA math_qa openbook_qa sciq medmcqa A 32,000 random subset (seed 42) textquestion-answering100K<n<1M0 likes509 downloads1y agoHugging Face08notpaulmartin /spider_mcqa_v0.2_full Spider-MCQA Converted Spider Text-to-SQL (Paper: Yu et al., 2018; HF Dataset) test set into multiple-choice. The dataset contains 1,034 examples. Dataset Fields Each JSON record contains: query: the schema and natural-language question prompt. gold_answer: the correct SQL answer. options: four SQL answer options, including the gold answer and three generated distractors. correct_option_index: the index of the correct answer in options. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/notpaulmartin/spider_mcqa_v0.2_full.textmultiple-choice1K<n<10K0 likes508 downloads4mo agoHugging Face09EleutherAI /wmdp_bio_robust_mcqatext1K<n<10K0 likes504 downloads1y agoHugging Face10andresnowak /MNLP_M3_mcqa_datasetThis dataset contains the MCQA and instruction finetuning datasets (and the test and validation splits are only used for testing not for training): The messages column is used by the instruction finetuning dataset The choices, question, context, and answer columns are used by the MCQA dataset For the MCQA dataset (of only single answer) contains a mixture of the train, validation and test splits from this datasets as to have for training and testing: mmlu auxiliary train we only use the… See the full description on the dataset page: https://huggingface.co/datasets/andresnowak/MNLP_M3_mcqa_dataset.text100K<n<1M0 likes377 downloads1y agoHugging Face11Wizard0504 /MNLP_M3_mcqa_datasettext10K<n<100K0 likes359 downloads1y agoHugging Face12Atnafu /Afri-MCQA Afri-MCQA: Multimodal Cultural Question Answering for African Languages Paper Overview Afri-MCQA is the first multilingual cultural question-answering benchmark covering 8k Q&A pairs across 16 African languages from 13 countries. The benchmark offers parallel English-African language Q&A pairs across text and speech modalities, entirely created by native speakers. Supported Tasks Visual Question Answering (VQA): Multiple-choice and open-ended QA… See the full description on the dataset page: https://huggingface.co/datasets/Atnafu/Afri-MCQA.audioimage-text-to-text10K<n<100K18 likes346 downloads3mo agoHugging Face13timarni /MNLP_M3_mcqa_datasettext100K<n<1M0 likes302 downloads1y agoHugging Face14thainamhoang /MNLP_M3_mcqa_datasettext100K<n<1M0 likes281 downloads1y agoHugging Face15aaronwzl /mcqa_calibration_datasettext1K<n<10K1 likes279 downloads2y agoHugging Face16jchang153 /copycolors_mcqa Synthetic copycolors_mcqa (4 answer choices) This dataset is a synthetic extension of mib-bench/copycolors_mcqa, restricted to the 4-choice setting used in this repository. It keeps only these counterfactual families: answerPosition_counterfactual randomLetter_counterfactual answerPosition_randomLetter_counterfactual The export uses a single train split. Each row contains one base prompt and one source row for each of the three counterfactual types, so the dataset is balanced… See the full description on the dataset page: https://huggingface.co/datasets/jchang153/copycolors_mcqa.textquestion-answering10K<n<100K0 likes251 downloads6mo agoHugging Face17andresnowak /MNLP_M2_mcqa_datasetThis dataset contains the MCQA and instruction finetuning datasets: The messages column is used by the instruction finetuning dataset The choices, question, context, and answer columns are used by the MCQA dataset For the MCQA dataset (of only single answer) contains a mixture of the train, validation and test splits from this datasets as to have for training and testing: mmlu auxiliary train we only use the stem subsets mmlu we only use the stem subsets ai2_arc ScienceQA math_qa… See the full description on the dataset page: https://huggingface.co/datasets/andresnowak/MNLP_M2_mcqa_dataset.textquestion-answering100K<n<1M0 likes250 downloads1y agoHugging Face18shulijia /MNLP_M3_mcqa_dataset_openbookqa_cottabular1K<n<10K0 likes234 downloads1y agoHugging Face19nvidia /Nemotron-RL-knowledge-web_search-mcqa Dataset Description: The Nemotron-RL-knowledge-web_search-mcqa is a multi-domain synthetic dataset designed to improve science and general reasoning in large language models (LLMs). It is a filtered subset of the OpenScienceReasoning-2 dataset and contains multiple-choice question–answer pairs spanning diverse domains: physics, biology, mathematics, humanities, computer science, engineering, chemistry, and others. This dataset is released as part of NVIDIA NeMo Gym, a framework… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-knowledge-web_search-mcqa.text1K<n<10K17 likes234 downloads17d agoHugging Face20open-athena /nemotron-gym-knowledge-web-search-mcqa-qwen3.5-122b-131k-opencode-traces Agent trace dataset Decoding the literal token IDs The prompt_token_ids / completion_token_ids / logprobs columns are the verbatim tokens the serving engine emitted, stored PER AGENT STEP as a list-of-lists (one inner list per turn). To turn them back into text you MUST use the exact tokenizer the model was served with — a generic same-family tokenizer will decode word tokens to garbage. Served model / tokenizer source: Qwen/Qwen3.5-122B-A10B-FP8 from transformers… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/nemotron-gym-knowledge-web-search-mcqa-qwen3.5-122b-131k-opencode-traces.text1K<n<10K0 likes222 downloads3mo agoHugging Face21shulijia /MNLP_M3_mcqa_dataset_qasc_cottext1K<n<10K0 likes213 downloads1y agoHugging Face22open-athena /nemotron-gym-knowledge-mcqa-qwen3.5-122b-32k-tracestext1K<n<10K0 likes200 downloads4mo agoHugging Face23rl-rag /drtulu_v2_nemotron_web_search_mcqatext1K<n<10K0 likes186 downloads8mo agoHugging Face24GingerBled /M3_MCQA_cs_science_mathtext100K<n<1M0 likes181 downloads1y agoHugging Face25thewordsmiths /stem_mcqatext10K<n<100K3 likes179 downloads2y agoHugging Face26lots-o /finance-law-mcqa 국가법령정보센터 문서를 기반으로 skt/A.X-4.0를 활용하여 생성 text10K<n<100K0 likes174 downloads1y agoHugging Face27stellaathena /math_mcqa MATH-MCQA A multiple choice adaptation of the MATH dataset containing 12,498 competition-level mathematics problems. Key Statistics Metric Value Total Examples 12,498 Train Split 7,498 Test Split 5,000 Categories 7 (Algebra, Intermediate Algebra, Prealgebra, Geometry, Number Theory, Counting & Probability, Precalculus) Difficulty Levels 5 (Level 1 = Easiest, Level 5 = Hardest/Competition-level) Format 4-option multiple choice (1 correct answer + 3… See the full description on the dataset page: https://huggingface.co/datasets/stellaathena/math_mcqa.text10K<n<100K2 likes163 downloads9mo agoHugging Face28stochastic-parrots /STEM-MCQA-Synthetic-55Ktext10K<n<100K0 likes157 downloads1y agoHugging Face29eve-esa /mcqa-multiple-answers Dataset Summary EVE-mcqa-multiple-answers is a Multiple-Choice Question Answering (MCQA) dataset designed to evaluate the performance of language models in the domain of Earth Observation (EO). The dataset consists of questions related to EO concepts, technologies, and applications, each accompanied by multiple answer choices, with one or more correct answer. Dataset Structure Each example in the dataset contains an arbitrary number of possible choices and one or more… See the full description on the dataset page: https://huggingface.co/datasets/eve-esa/mcqa-multiple-answers.textmultiple-choicen<1K0 likes157 downloads6mo agoHugging Face30eve-esa /mcqa-single-answer Dataset Summary EVE-mcqa-single-answer is a Multiple-Choice Question Answering (MCQA) dataset designed to evaluate the performance of language models in the domain of Earth Observation (EO). The dataset consists of questions related to EO concepts, technologies, and applications, each accompanied by multiple answer choices with exactly one correct answer. Unlike multi-answer MCQA datasets, each question in this dataset has only a single correct choice, making it suitable for… See the full description on the dataset page: https://huggingface.co/datasets/eve-esa/mcqa-single-answer.textmultiple-choice1K<n<10K0 likes155 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.