agentic-reasoning
gemma-4-E4B-Agentic-Sol-Fable-Reasoning-GeminiCLI-GGUFgemma-4-E4B-Agentic-Opus-Reasoning-GeminiCLI-GGUFgemma-4-E4B-Agentic-Opus-Reasoning-GeminiCLI-mlx-4bitgemma-4-E4B-Agentic-Sol-Fable-Reasoning-GeminiCLI-mlx-4bitgemma-4-E4B-Agentic-Opus-Reasoning-GeminiCLI-mlx-4bitgemma-4-E4B-Agentic-Opus-Reasoning-GeminiCLI-GGUFLlama-3.1-8B-Agentic-ReasoningxDAN-L1-Edge-Agentic-7b-reasoning-kd-e1
Agentic-Diagnostic-Reasoning-with-Multimodal-SLMs-via-Reinforcement-Learningdaily-paper-2026-09-11-reasoning-verbosity-tax-agentic-cost
The Reasoning Verbosity Tax: Per-Turn Cost-Quality Frontiers of Explicit vs. Compact Chain-of-Thought in Self-Hosted Agentic LLMs on H200
TL;DR — An analytical paper that turns the Qwen3 thinking toggle into a priced dial for self-hosted agentic LLM loops: explicit chain-of-thought persisted in context compounds a per-turn verbosity tax quadratically in the horizon (Theta(T^2)) versus linearly (Theta(T)) under a discard policy, the relative tax widens monotonically toward a… See the full description on the dataset page: https://huggingface.co/datasets/thaki-AI/daily-paper-2026-09-11-reasoning-verbosity-tax-agentic-cost.agentic-reasoning-benchmark
Agentic & Reasoning Benchmark (ARB) – Expanded
Ein synthetischer Benchmark mit 2.550 Fragen und Lösungen, optimiert für die Evaluation von Agentic Capabilities und Reasoning.
Überblick
Eigenschaft
Wert
Anzahl Beispiele
2.550
Kategorien
8
Schwierigkeitsgrade
easy / medium / hard
Formate
CSV + JSON
Reproduzierbarkeit
Generator-Skript (seed=42) enthalten
Lizenz
CC-BY-4.0
Kategorien
Kategorie
Anzahl
Beschreibung… See the full description on the dataset page: https://huggingface.co/datasets/roskosmos19/agentic-reasoning-benchmark.agentic-reasoning-trace-summaries-40k
Reasoning Summary JSON Dataset
We built this dataset to train models to turn long reasoning/work traces into short structured summaries.
Each example has a verbose trace in input and a compact JSON summary in output. The summary is shaped like the kind of progress update we want a model to produce while it is working: a title, a subtitle, a short summary, and the current task.
The dataset is JSONL with 43,734 rows. The rows are ordered from longest to shortest so long-context… See the full description on the dataset page: https://huggingface.co/datasets/mithulaartigala/agentic-reasoning-trace-summaries-40k.tool-reasoning-sft-CODING-nvidia-Nemotron-Agentic-v1
Nemotron-Agentic-v1 — Cleaned & Rectified
335k multi-turn agentic tool-use trajectories from NVIDIA's Nemotron-Agentic-v1, converted into a strict reasoning + tool-call format with validated FSM transitions.
Origin
Derived from nvidia/Nemotron-Agentic-v1.
Nemotron-Agentic-v1 is a synthetic dataset of multi-turn conversations where language models decompose user goals, decide when to call tools, and reason over tool outputs. Trajectories are generated by simulating user… See the full description on the dataset page: https://huggingface.co/datasets/AmanPriyanshu/tool-reasoning-sft-CODING-nvidia-Nemotron-Agentic-v1.mimo-v25-agentic-reasoning-traces
Mimo V2.5 Agentic Reasoning Traces
Repository: OusiaResearch/mimo-v25-agentic-reasoning-traces
Paper: ( forthcoming )
License: CC BY-NC 4.0
Overview
A 14,726-row corpora of agentic reasoning traces generated via Xiaomi Mimo V2.5 Pro, structured for fine-tuning and evaluating language models on consequential, real-world decision-making tasks.
The corpus has two components:
Component
Rows
Description
mimo_v25_reasoning_traces.jsonl
9,287
General reasoning traces… See the full description on the dataset page: https://huggingface.co/datasets/OusiaResearch/mimo-v25-agentic-reasoning-traces.
