Team Ai
29 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01haritzpuerto /controlling-reasoning-models-privacy-outputs Model Outputs Dataset Card Go to **Files and versions** tab to access the data. Dataset Description This dataset contains the raw model generations (reasoning traces and final answers) produced in the experiments described in our paper Controllable Reasoning Models are Private Thinkers. It aggregates outputs for: two model families: Qwen 3 and Phi 4, multiple model sizes (1.7B–14B), five variants per model (baseline, RT-IF–optimized, FA-IF–optimized… See the full description on the dataset page: https://huggingface.co/datasets/haritzpuerto/controlling-reasoning-models-privacy-outputs.0 likes247 downloads1mo agoHugging Face02hanspeterlyngsoeraaschoujensen /reasoning-models-noncritical-artifacts0 likes178 downloads6mo agoHugging Face03Dongwei /reasoning_world_modeltext10K<n<100K12 likes142 downloads2y agoHugging Face04OzTianlu /Why_Reasoning_Models_Collapse_Themselves_in_Reasoning Why Reasoning Models Collapse Themselves in Reasoning 4 Algorithmic Atoms Reveal the Geometric Truth Abstract This paper presents LeftAndRight, a diagnostic framework using four algorithmic primitives (>>, <<, 1, 0) to reveal a fundamental property of transformer representations: they geometrically collapse backward operations, regardless of attention architecture. Key Discovery The counterintuitive finding: We initially hypothesized that causal attention masks… See the full description on the dataset page: https://huggingface.co/datasets/OzTianlu/Why_Reasoning_Models_Collapse_Themselves_in_Reasoning.documentn<1K0 likes117 downloads11mo agoHugging Face05lgyeee /where-larger-models-excel-reasoning-traces0 likes117 downloads6mo agoHugging Face06OzTianlu /A_Reasoning_Critique_of_Diffusion_Models A Reasoning Critique of Diffusion Models Author: Zixi "Oz" Li (李籽溪) Date: December 12, 2025 Type: Theoretical AI Research (Geometry, Reasoning Theory) Citation @misc{oz_lee_2025, author = { Oz Lee }, title = { A_Reasoning_Critique_of_Diffusion_Models (Revision 267326d) }, year = 2025, url = { https://huggingface.co/datasets/OzTianlu/A_Reasoning_Critique_of_Diffusion_Models }, doi = { 10.57967/hf/7243 }… See the full description on the dataset page: https://huggingface.co/datasets/OzTianlu/A_Reasoning_Critique_of_Diffusion_Models.documentn<1K3 likes115 downloads10mo agoHugging Face07LLM-OS-Models /LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-Raw LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-Raw Stage2 raw LFM chat JSONL shards: Korean domain, behavior, SWE/coding, reasoning, finance, legal, Text2SQL. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-Raw.0 likes90 downloads3mo agoHugging Face08modelscope /MMMU-Reasoning-Distill-Validation中文版本 Description MMMU-Reasoning-Distill-Validation is a Multi-Modal reasoning dataset that contains 839 image descriptions and natural language inference data samples. This dataset is built upon the validation set of MMMU. The construction process begins with using Qwen2.5-VL-72B-Instruct for image understanding and generating detailed image descriptions, followed by generating reasoning conversations using the DeepSeek-R1 model. Its main features are as follows: Use the… See the full description on the dataset page: https://huggingface.co/datasets/modelscope/MMMU-Reasoning-Distill-Validation.imagen<1K2 likes75 downloads2y agoHugging Face09OzTianlu /Semigroup_Reasoning_Model_A_Scalpel Semigroup Reasoning Model: A Scalpel Formalizing Sparse Neural Circuits as Reasoning Dynamics 🎯 Central Question How do we formalize the interpretability of reasoning processes? This work establishes reasoning as a semigroup dynamical system, providing the first formal equivalence between sparse neural circuits and algebraic reasoning dynamics. We prove that: Reasoning is a semigroup orbit problem, not a vector space embedding task. 🔬 Key Contributions… See the full description on the dataset page: https://huggingface.co/datasets/OzTianlu/Semigroup_Reasoning_Model_A_Scalpel.documentn<1K1 likes51 downloads10mo agoHugging Face10exp-models /koen-reasoning-calibration-v2text1K<n<10K1 likes45 downloads2y agoHugging Face11ziadrone /ida-reasoning-model IDA Reasoning Model This model was trained using Imitation, Distillation, and Amplification (IDA) on multiple reasoning datasets. Training Details Teacher Model: deepseek-ai/DeepSeek-R1-Distill-Qwen-7B Student Model: Qwen/Qwen3-1.7B Datasets: 4 reasoning datasets Total Samples: 600 Training Method: IDA (Iterative Distillation and Amplification) Datasets Used gsm8k HuggingFaceH4/MATH-500 MuskumPillerum/General-Knowledge SAGI-1/reasoningData_200k… See the full description on the dataset page: https://huggingface.co/datasets/ziadrone/ida-reasoning-model.textn<1K0 likes38 downloads1y agoHugging Face12jaygala24 /reasoning-models-interpretability-artifacts Reasoning Models Interpretability Artifacts This dataset contains intermediate artifacts for studying reasoning traces in open-weight language models. It includes annotated-trace hidden representations and spectral metrics computed over reasoning-step categories. The artifacts are intended for analysis and sharing, not for direct datasets.load_dataset(...) loading as a tabular dataset. Contents annotated_traces_reprs/ <model>/ config.json index.json… See the full description on the dataset page: https://huggingface.co/datasets/jaygala24/reasoning-models-interpretability-artifacts.1K<n<10K0 likes36 downloads6mo agoHugging Face13ziadrone /ida-reasoning-model1textn<1K0 likes32 downloads1y agoHugging Face14dvilasuero /jailbreak-classification-reasoning-modelstextn<1K0 likes27 downloads2y agoHugging Face15mramazan /nvidia-nemotron-model-reasoning-dataset-turkish Nemotron Reasoning Challenge - Turkish Turkish translation of the training data from NVIDIA's Nemotron Model Reasoning Challenge Each row is a reasoning puzzle framed in an "Alice's Wonderland" setting. Given a few input/output examples, the model needs to figure out the hidden rule and apply it to a new input. Category Rows Description bit 1602 Hidden bit manipulation rule on 8-bit binary numbers grav 1597 Falling distance with a modified gravitational constant… See the full description on the dataset page: https://huggingface.co/datasets/mramazan/nvidia-nemotron-model-reasoning-dataset-turkish.texttext-generation1K<n<10K1 likes24 downloads4mo agoHugging Face16LLM-OS-Models /LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-4K LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-4K Stage2 diverse Korean/SWE/reasoning prepared SFT arrays, LFM tokenizer. This dataset is part of the LFM2.5-8B-A1B-KO-SFT / Agentic SFT workflow. Main SFT model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-SFT CPT base model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-CPT-FULL Agentic follow-up model: https://huggingface.co/LLM-OS-Models/LFM2.5-8B-A1B-KO-Agentic-SFT SFT GitHub:… See the full description on the dataset page: https://huggingface.co/datasets/LLM-OS-Models/LFM2.5-KO-SFT-Stage2-Diverse-KoSWE-Reasoning-LFMChat-4K.0 likes24 downloads3mo agoHugging Face17jasonkung98 /NVIDIA-Nemotron-Model-Reasoning-Challengetext1K<n<10K1 likes22 downloads7mo agoHugging Face18ariefansclub /humanoid-world-model-reasoning Humanoid World Model Reasoning Dataset for training internal world models and reasoning loops in humanoid AI. textn<1K0 likes18 downloads9mo agoHugging Face19booba-uz /test-reasoning-model-datasettext10K<n<100K0 likes16 downloads2y agoHugging Face20letridung07 /Medical-Knowledge-Benchmark-for-Reasoning-AI-ModelsLink to the original dataset on HuggingFace: https://huggingface.co/datasets/FreedomIntelligence/medical-o1-reasoning-SFT We are going to test on 4 free Reasoning AI models: Kimi K1.5-extended-thinking Deepseek R1 Qwen3-235B-A22B Gemini 2.5 Pro Settings: Web Search: Disabled❌ Reasoning: Enabled✅ Temperature: Default 0 likes16 downloads1y agoHugging Face21exp-models /korean-reasoning-mixture-20250203-previewgatedtext10K<n<100K6 likes11 downloads2y agoHugging Face22thia-co /Core-Model-SFT-Adv-Reasoningtabularn<1K0 likes10 downloads1y agoHugging Face23kavyachouhan /reasoning-enhanced-image-gen-modelimage1K<n<10K0 likes10 downloads8mo agoHugging Face24Maitreyajayaraj /quantitative_modeling_reasoning_v2textn<1K0 likes9 downloads7mo agoHugging Face25exp-models /koen-reasoning-calibration-v1text1K<n<10K0 likes7 downloads2y agoHugging Face26Maitreyajayaraj /scientific_modeling_reasoning0 likes6 downloads7mo agoHugging Face27Maitreyajayaraj /advanced_mathematical_modelling_reasoning_v1textn<1K0 likes6 downloads7mo agoHugging Face28Katakalyst /reasoning-model-kb0 likes6 downloads5mo agoHugging Face29exp-models /korean-reasoning-mixture-20250203-preview-calibrationtext1K<n<10K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.