Team Ai
15 results

answer-only

dougalldeepmind /2026-10-05-odcv-qwen36-0-da-15-answer-only odcv eval of dougalldeepmind/2026-10-05-qwen36-0-da-15-answer-only (mode=think) field value experiment odcv eval of dougalldeepmind/2026-10-05-qwen36-0-da-15-answer-only (mode=think) date_generated 2026-10-05 constitution none source_repo teaching_claude_why_replication @ 0697412633c24c8ab5cb9d1c976c1f22cb7c107a models {"target": "dougalldeepmind/2026-10-05-qwen36-0-da-15-answer-only", "target_revision": "291e04a399ba12c20840ed876ff4967f61f79fec", "base":… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-10-05-odcv-qwen36-0-da-15-answer-only.0 likes509 downloads5d agoHugging Faceskrishna /gsm8k_only_answerThe data is exactly like the original GSM8k (https://huggingface.co/datasets/gsm8k ), but with the label consisting of the correct answer(one number) only. @misc{krishna2024gsmansweronly, title={GSM8k (Answer only)}, author={Satyapriya Krishna}, year={2023}, url={skrishna/gsm8k_only_answer}, } text1K<n<10K2 likes453 downloads2y agoHugging Faceaddy88 /nq-question-answeronlytext100K<n<1M1 likes235 downloads5y agoHugging Facehomerquan /boardgamebench-answer-only BoardGameBench Answer-Only Reasoning Dataset This dataset contains 1,282,766 board-game reasoning examples generated from BoardGameBench, a benchmark and data-generation project for evaluating language models on structured board-game decision making. Each row asks a model to inspect a legal board position and return the best move. The format is intentionally simple: id,prompt,answer The answer field is the target move label, such as C4, f6, 11,7, or e2-d3. This makes the dataset… See the full description on the dataset page: https://huggingface.co/datasets/homerquan/boardgamebench-answer-only.text-generation1M<n<10M0 likes165 downloads5mo agoHugging Faceweikaih /imaginative-perception-token-mvc-answeronly Citation Released with the paper Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models (arXiv:2606.03988): @misc{bigverdi2026imaginativeperceptiontokensenhance, title={Imaginative Perception Tokens Enhance Spatial Reasoning in Multimodal Language Models}, author={Mahtab Bigverdi and Linjie Li and Weikai Huang and Yiming Liu and Jaemin Cho and Jieyu Zhang and Tuhin Kundu and Chris Dangjoo Kim and Zelun Luo and Linda Shapiro and Ranjay… See the full description on the dataset page: https://huggingface.co/datasets/weikaih/imaginative-perception-token-mvc-answeronly.text10K<n<100K0 likes143 downloads4mo agoHugging Facedougalldeepmind /2026-09-16-odcv-qwen36-0-da-7-answer-only odcv eval of dougalldeepmind/2026-09-16-qwen36-0-da-7-answer-only (mode=think) field value experiment odcv eval of dougalldeepmind/2026-09-16-qwen36-0-da-7-answer-only (mode=think) date_generated 2026-09-16 constitution none source_repo teaching_claude_why_replication @ 0ee0e1248d478f976806b783deb0987c04cbaa5c models {"target": "dougalldeepmind/2026-09-16-qwen36-0-da-7-answer-only", "target_revision": "52adc308c378457a94b2eb900802dc424ab5540f", "base":… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-09-16-odcv-qwen36-0-da-7-answer-only.0 likes125 downloads18d agoHugging Face