Team Ai
20 results

zebra

multimodal-reasoning-lab /Zebra-CoT Zebra‑CoT A diverse large-scale dataset for interleaved vision‑language reasoning traces. Dataset Description Zebra‑CoT is a diverse large‑scale dataset with 182,384 samples containing logically coherent interleaved text‑image reasoning traces across four major categories: scientific reasoning, 2D visual reasoning, 3D visual reasoning, and visual logic & strategic games. Dataset Structure Each example in Zebra‑CoT consists of: Problem statement:… See the full description on the dataset page: https://huggingface.co/datasets/multimodal-reasoning-lab/Zebra-CoT.imageany-to-any100K<n<1M80 likes12k downloads8mo agoHugging FaceWildEval /ZebraLogicPaper: https://huggingface.co/papers/2502.01100 Arxiv: https://arxiv.org/abs/2502.01100 Citation @article{zebralogic2025, title={ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning}, author={Bill Yuchen Lin and Ronan Le Bras and Kyle Richardson and Ashish Sabharwal and Radha Poovendran and Peter Clark and Yejin Choi}, year={2025}, url={https://arxiv.org/abs/2502.01100}, } @article{dziri2024faith, title={Faith and fate: Limits of transformers on… See the full description on the dataset page: https://huggingface.co/datasets/WildEval/ZebraLogic.text1K<n<10K18 likes5.7k downloads2y agoHugging Faceleafspark /OpenRouter-ZebraLogicBench OpenRouter-ZebraLogicBench This repository contains a single Python file evaluation script for the allenai/ZebraLogicBench dataset. The script is adapted from ZeroEval and can be used to evaluate language models on logical reasoning tasks. Key Features Single file implementation for easy use Compatible with OpenAI-like APIs (base URL can be modified in eval_zebra.py) Example results provided for Claude 3 Haiku Usage Requirements Access to the private dataset:… See the full description on the dataset page: https://huggingface.co/datasets/leafspark/OpenRouter-ZebraLogicBench.question-answering1K<n<10K2 likes1.6k downloads2y agoHugging FaceSightLinks /YOLO-OBB-Zebra-Crossings-Datasetimage1K<n<10K0 likes1.3k downloads2y agoHugging FaceRuoliuYang /textlatent_zebra_thinkmorph_armAB Text-Latent (Arm A) vs All-Latent (Arm B) — Zebra-CoT + ThinkMorph 35638 samples/arm, 18 categories. Schema = ULVR/williamium style (sample_id, category, source_dataset, question, answer, input_image, intermediate_image_N, num_intermediate_steps, messages_json). armA_text_latent: real decoded text CoT + latent visual blocks (intermediate_image_1..3). armB_render_latent: reasoning text RENDERED to images, all-latent baseline (intermediate_image_1..17). messages_json = full Monet… See the full description on the dataset page: https://huggingface.co/datasets/RuoliuYang/textlatent_zebra_thinkmorph_armAB.image100K<n<1M0 likes1.2k downloads3mo agoHugging Facealexandrainst /multi-zebra-logic Dataset Card for the MultiZebraLogic dataset This dataset includes zebra puzzles in 39 European and 5 non-European languages and in two sizes: 2x3 and 4x5. It can be used for evaluating logical reasoning ability. The data has been generated using the code in this repo. Dataset Details Dataset Description Zebra puzzles are a type of constraint satisfaction problem. They describe a number of objects, N_objects, that each have attributes… See the full description on the dataset page: https://huggingface.co/datasets/alexandrainst/multi-zebra-logic.texttext-generation100K<n<1M1 likes850 downloads2mo agoHugging Face