Team Ai
20 results

Thinking

zhiyuanhucs /nemotron-student-fail-v41-clean-thinking DeepSeek-V4.1 clean and action-only trajectories with Nemotron outcomes DeepSeek-V4.1 reward-1 trajectories rebuilt from the complete teacher audit under v57-test-path-component-boundary+v57-target-source-recheck. The V4.1 reward and trajectory tier do not by themselves prove that Nemotron failed. Student outcomes are joined from nemotron-prolike-coverage-audit-20261001.json. A student failure requires either complete required-test results with reward 0, or an individually… See the full description on the dataset page: https://huggingface.co/datasets/zhiyuanhucs/nemotron-student-fail-v41-clean-thinking.tabulartext-generationn<1K1 likes13k downloads7d agoHugging Facea-m-team /AM-Thinking-v1-Distilled 📘 Dataset Summary AM-Thinking-v1 and Qwen3-235B-A22B are two reasoning datasets distilled from state-of-the-art teacher models. Each dataset contains high-quality, automatically verified responses generated from a shared set of 1.89 million queries spanning a wide range of reasoning domains. The datasets share the same format and verification pipeline, allowing for direct comparison and seamless integration into downstream tasks. They are intended to support the development of… See the full description on the dataset page: https://huggingface.co/datasets/a-m-team/AM-Thinking-v1-Distilled.text-generation1M<n<10M64 likes12k downloads1y agoHugging FaceHuggingFaceH4 /Multilingual-Thinking Dataset summary Multilingual-Thinking is a reasoning dataset where the chain-of-thought has been translated from English into one of 4 languages: Spanish, French, Italian, and German. The dataset was created by sampling 1k training samples from the SystemChat subset of SmolTalk2 and translating the reasoning traces with another language model. This dataset was used in the OpenAI Cookbook to fine-tune the OpenAI gpt-oss models. You can load the dataset using: from datasets import… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceH4/Multilingual-Thinking.texttext-generation1K<n<10K120 likes8.9k downloads1y agoHugging FaceLoneResearch /thinking-model-activations0 likes3.4k downloads8mo agoHugging FaceShareLab-SII /thinking_droid_lerobot_output_qwen3vlimage1M<n<10M0 likes3.4k downloads6mo agoHugging Facellm-jp /llm-jp-4.1-thinking-sft-data llm-jp-4.1-thinking-sft-data Overview This dataset is a supervised fine-tuning (SFT) dataset used to train llm-jp-4.1-*-thinking models. This dataset is constructed from prompts and conversations collected from multiple data sources. For most subsets, reasoning processes and final responses used for LLM-jp-4.1 SFT were generated or augmented using gpt-oss-120b. For the tool-calling and agentic data derived from NVIDIA Nemotron datasets, the original conversations… See the full description on the dataset page: https://huggingface.co/datasets/llm-jp/llm-jp-4.1-thinking-sft-data.text1M<n<10M4 likes2.8k downloads16d agoHugging Face