Team Ai
16 results

agentic-reasoning

n1ghtf4l1 /Agentic-Diagnostic-Reasoning-with-Multimodal-SLMs-via-Reinforcement-Learning15 likes175 downloads11mo agoHugging Facethaki-AI /daily-paper-2026-09-11-reasoning-verbosity-tax-agentic-cost The Reasoning Verbosity Tax: Per-Turn Cost-Quality Frontiers of Explicit vs. Compact Chain-of-Thought in Self-Hosted Agentic LLMs on H200 TL;DR — An analytical paper that turns the Qwen3 thinking toggle into a priced dial for self-hosted agentic LLM loops: explicit chain-of-thought persisted in context compounds a per-turn verbosity tax quadratically in the horizon (Theta(T^2)) versus linearly (Theta(T)) under a discard policy, the relative tax widens monotonically toward a… See the full description on the dataset page: https://huggingface.co/datasets/thaki-AI/daily-paper-2026-09-11-reasoning-verbosity-tax-agentic-cost.0 likes161 downloads1mo agoHugging Faceroskosmos19 /agentic-reasoning-benchmark Agentic & Reasoning Benchmark (ARB) – Expanded Ein synthetischer Benchmark mit 2.550 Fragen und Lösungen, optimiert für die Evaluation von Agentic Capabilities und Reasoning. Überblick Eigenschaft Wert Anzahl Beispiele 2.550 Kategorien 8 Schwierigkeitsgrade easy / medium / hard Formate CSV + JSON Reproduzierbarkeit Generator-Skript (seed=42) enthalten Lizenz CC-BY-4.0 Kategorien Kategorie Anzahl Beschreibung… See the full description on the dataset page: https://huggingface.co/datasets/roskosmos19/agentic-reasoning-benchmark.textquestion-answering1K<n<10K1 likes78 downloads1mo agoHugging Facemithulaartigala /agentic-reasoning-trace-summaries-40k Reasoning Summary JSON Dataset We built this dataset to train models to turn long reasoning/work traces into short structured summaries. Each example has a verbose trace in input and a compact JSON summary in output. The summary is shaped like the kind of progress update we want a model to produce while it is working: a title, a subtitle, a short summary, and the current task. The dataset is JSONL with 43,734 rows. The rows are ordered from longest to shortest so long-context… See the full description on the dataset page: https://huggingface.co/datasets/mithulaartigala/agentic-reasoning-trace-summaries-40k.texttext-generation10K<n<100K10 likes73 downloads2mo agoHugging FaceAmanPriyanshu /tool-reasoning-sft-CODING-nvidia-Nemotron-Agentic-v1 Nemotron-Agentic-v1 — Cleaned & Rectified 335k multi-turn agentic tool-use trajectories from NVIDIA's Nemotron-Agentic-v1, converted into a strict reasoning + tool-call format with validated FSM transitions. Origin Derived from nvidia/Nemotron-Agentic-v1. Nemotron-Agentic-v1 is a synthetic dataset of multi-turn conversations where language models decompose user goals, decide when to call tools, and reason over tool outputs. Trajectories are generated by simulating user… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-CODING-nvidia-Nemotron-Agentic-v1.texttext-generation100K<n<1M1 likes61 downloads7mo agoHugging FaceOusiaResearch /mimo-v25-agentic-reasoning-traces Mimo V2.5 Agentic Reasoning Traces Repository: OusiaResearch/mimo-v25-agentic-reasoning-traces Paper: ( forthcoming ) License: CC BY-NC 4.0 Overview A 14,726-row corpora of agentic reasoning traces generated via Xiaomi Mimo V2.5 Pro, structured for fine-tuning and evaluating language models on consequential, real-world decision-making tasks. The corpus has two components: Component Rows Description mimo_v25_reasoning_traces.jsonl 9,287 General reasoning traces… See the full description on the dataset page: https://huggingface.co/datasets/OusiaResearch/mimo-v25-agentic-reasoning-traces.n<1K4 likes50 downloads5mo agoHugging Face