Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AgenticFinLab /PortBench-Market PortBench Market Base Dataset Dataset Description A ten-year (Jan 2015–Dec 2025) daily financial dataset covering 183 instruments across six heterogeneous asset classes, designed for multi-asset portfolio management research and LLM evaluation. Asset Coverage Asset Class Instruments Data Fields Sources Equities 126 OHLCV + return Yahoo Finance (ETFs: broad market, sector, factor, international) Bonds 16 Close + return (ETFs); yield… See the full description on the dataset page: https://huggingface.co/datasets/AgenticFinLab/PortBench-Market.tabulartime-series-forecasting1K<n<10K3 likes243 downloads4mo agoHugging Face02agentic-moral-alignment /matrix-game-evaltabular10K<n<100K0 likes225 downloads5mo agoHugging Face03cx-cmu /deepresearchgym-agentic-search-logs DeepResearchGym Agentic Search Logs This repository hosts the dataset accompanying the paper “Agentic Search in the Wild” (arXiv: https://arxiv.org/abs/2601.17617). The dataset contains 14M+ search queries collected via DeepResearchGym (DRGym), an open-source search API designed for DeepResearch-style agentic search. For more background on DRGym, see: https://arxiv.org/abs/2505.19253. All records have been anonymized and shuffled to prevent re-identification, and we additionally… See the full description on the dataset page: https://huggingface.co/datasets/cx-cmu/deepresearchgym-agentic-search-logs.tabulartext-retrieval10M<n<100M16 likes214 downloads8mo agoHugging Face04AgenticFinLab /PyFi-600K Dataset Card for PyFi-600K This dataset card aims to be a introduction for PyFi-600K, A financial VLM dataset containing 600K question-answer pairs generated via Adversarial agents. AgenticFinLab/PyFi-600K/ ├── README.md # Dataset documentation and description ├── images.zip # Compressed image files ├── PyFi-600K-dataset.csv # Q&A pairs in CSV format ├── PyFi-600K-dataset.json # Q&A pairs in JSON format ├── PyFi-600K-chain-dataset.json # Chain of Thought Q&A pairs dataset └──… See the full description on the dataset page: https://huggingface.co/datasets/AgenticFinLab/PyFi-600K.imagequestion-answering100K<n<1M1 likes167 downloads10mo agoHugging Face05agentic-moral-alignment /persona-and-other-evals Qwen3.5-9B AMA adapters — persona evals Inference code, the data it produced, and the tools that turn that data into tables and an HTML viewer. The evals are Anthropic's persona set, scored in three regimes: teacher-forced logprob of the answer literal, greedy answer with the reasoning block pre-closed, and a full 16k-budget reasoning trace. Pinned models base unsloth/Qwen3.5-9B @ 005429cee5cb648998cf2b70eebdd83175989c9a util… See the full description on the dataset page: https://huggingface.co/datasets/agentic-moral-alignment/persona-and-other-evals.tabularn<1K0 likes164 downloads19d agoHugging Face06skandotai /agentic-ontology-of-work Agentic Ontology of Work (AOW) The Agentic Ontology of Work (AOW) is an open model for describing work done by AI agents, people, and systems in an enterprise. It defines each part of the work, the information each part carries, and how the parts connect. This dataset contains version 2.0.0 of the ontology, its validation files, and validated example data. Website: https://skandotai.github.io/agentic-ontology-of-work/ Source repository:… See the full description on the dataset page: https://huggingface.co/datasets/skandotai/agentic-ontology-of-work.textn<1K0 likes162 downloads16d agoHugging Face07agentic-learning-ai-lab /daily-oracle Daily Oracle 📰 Project Website📝 Paper - Are LLMs Prescient? A Continuous Evaluation using Daily News as the Oracle Daily Oracle is a continuous evaluation benchmark using automatically generated QA pairs from daily news to assess how the future prediction capabilities of LLMs evolve over time. Dataset Details Question Type: True/False (TF) & Multiple Choice (MC) Current Version* Time Span: 2020.01.01 - 2026.07.18 Size: 20,376 TF questions and 18,557 MC… See the full description on the dataset page: https://huggingface.co/datasets/agentic-learning-ai-lab/daily-oracle.textquestion-answering10K<n<100K4 likes130 downloads3mo agoHugging Face08witcheer /agentic-score-leaderboard 🛠️ Agentic Score Leaderboard — one RTX 5090 How well do local models actually drive a tool-using agent loop? Not single-call function-calling benchmarks — a real loop: native OpenAI tool-calling through llama-server, multi-step deterministic tasks, programmatic verification. Everything runs on a single RTX 5090 32GB. Updated 2026-06-17 · llama.cpp b9562 · --jinja native tool-calling · temp 0. Leaderboard # model params Agentic Score success tool-eff… See the full description on the dataset page: https://huggingface.co/datasets/witcheer/agentic-score-leaderboard.tabularn<1K3 likes126 downloads3mo agoHugging Face09dsoosai /agentic-enterprise-demo Agentic Enterprise demo data Fictional data behind the Agentforce, Claude and Snowflake demo in github.com/dsoosai/crm-scoring-agents (agentforce/ and snowflake/agentic/). The walkthrough is on the CRM Scoring Agents Space. A data-platform vendor sells capacity contracts to twelve customer accounts. The featured account, Omega Inc., is on pace to run out of contracted credits five weeks before its renewal, its AI workload is up sharply and its data science workload is falling… See the full description on the dataset page: https://huggingface.co/datasets/dsoosai/agentic-enterprise-demo.tabular10K<n<100K0 likes97 downloads15d agoHugging Face10agenticx /DrugbankVocabularytext10K<n<100K0 likes89 downloads1y agoHugging Face11Agentic-Payment /agentic-payments-readiness Agentic AI Payments Readiness Report 2026 — by Pink Agentic AI Payments Published by Pink Agentic AI Payments (by PinkWallet) — the approval layer between AI agents and company money: plain-language rules, per-agent budgets and human approvals decide each payment before it executes. Agents connect via MCP or REST. Try the free public sandbox: https://agentic-sandbox.pinkwallet.com (test credentials, no real money moves) · Product: https://pinkwallet.com/agentic/ · Examples:… See the full description on the dataset page: https://huggingface.co/datasets/Agentic-Payment/agentic-payments-readiness.texttabular-classificationn<1K0 likes82 downloads8d agoHugging Face12roskosmos19 /agentic-reasoning-benchmark Agentic & Reasoning Benchmark (ARB) – Expanded Ein synthetischer Benchmark mit 2.550 Fragen und Lösungen, optimiert für die Evaluation von Agentic Capabilities und Reasoning. Überblick Eigenschaft Wert Anzahl Beispiele 2.550 Kategorien 8 Schwierigkeitsgrade easy / medium / hard Formate CSV + JSON Reproduzierbarkeit Generator-Skript (seed=42) enthalten Lizenz CC-BY-4.0 Kategorien Kategorie Anzahl Beschreibung… See the full description on the dataset page: https://huggingface.co/datasets/roskosmos19/agentic-reasoning-benchmark.textquestion-answering1K<n<10K1 likes78 downloads1mo agoHugging Face13Agentic-Payment /agent-spending-controls-crosswalk Agentic AI Payments: Agent Spending Controls Crosswalk (2026) — by Pink Agentic AI Payments Published by Pink Agentic AI Payments (by PinkWallet) — the approval layer between AI agents and company money: plain-language rules, per-agent budgets and human approvals decide each payment before it executes. Agents connect via MCP or REST. Try the free public sandbox: https://agentic-sandbox.pinkwallet.com (test credentials, no real money moves) · Product:… See the full description on the dataset page: https://huggingface.co/datasets/Agentic-Payment/agent-spending-controls-crosswalk.texttabular-classificationn<1K0 likes74 downloads8d agoHugging Face14Finance-Agentic-AI /Financial-Reportstabular10K<n<100K0 likes69 downloads10mo agoHugging Face15agenticx /DILIranktext1K<n<10K0 likes64 downloads1y agoHugging Face16hug-the-trees /0skeng-agentic-web-benchmark 0skeng Index: synthetic UK collectibles prices The prices in this dataset are generated, not collected. A script produces every figure from a fixed seed. No marketplace was read, no listing was matched, and no real sale lies behind any number here. Do not use it to value, buy or sell anything, and do not repeat any figure in it as a market fact. What it is good for is testing: it has the shape, spread and internal consistency of a real weekly price index, in formats a program or… See the full description on the dataset page: https://huggingface.co/datasets/hug-the-trees/0skeng-agentic-web-benchmark.tabularn<1K1 likes53 downloads3d agoHugging Face17Finance-Agentic-AI /Financial-Advisory-Clientstabular1K<n<10K1 likes50 downloads7mo agoHugging Face18agenticx /DIQTAtextn<1K0 likes48 downloads1y agoHugging Face19jamesdborin /Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1-prompt-only Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts.… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1-prompt-only.tabular10K<n<100K0 likes46 downloads3mo agoHugging Face20jamesdborin /Nemotron-RL-Agentic-Function-Calling-Pivot-v1-prompt-only Nemotron-RL-Agentic-Function-Calling-Pivot-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Agentic-Function-Calling-Pivot-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts.… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Agentic-Function-Calling-Pivot-v1-prompt-only.tabular1K<n<10K0 likes45 downloads3mo agoHugging Face21wmaulanaaishq /indonesian-agentic-kyc-aml-dataset Indonesian Agentic AI KYC & AML Dataset Dataset Description This dataset is specifically curated for building Agentic AI systems in the Indonesian financial sector. It focuses on Know Your Customer (KYC), Anti-Money Laundering (AML), and Financial Fraud Detection. The data has been scraped from official Indonesian regulatory bodies (OJK, Bank Indonesia) and top-tier financial news portals, then processed through an ETL pipeline to generate enriched RAG… See the full description on the dataset page: https://huggingface.co/datasets/wmaulanaaishq/indonesian-agentic-kyc-aml-dataset.tabularn<1K1 likes44 downloads2mo agoHugging Face22ravishgupta /ate-agentic-coverage O*NET Agentic Coverage Dataset A task-level estimate, for all 18,796 O*NET work tasks — the full O*NET task corpus, economy-wide — of how much of each task a current autonomous AI agent could complete end to end, with no human in the loop. This dataset supports task-level research on AI exposure and human-AI task allocation: which tasks an agent can already finish alone, which need a human for part of the work, and which still require a human throughout. Built as a companion… See the full description on the dataset page: https://huggingface.co/datasets/ravishgupta/ate-agentic-coverage.tabular10K<n<100K1 likes44 downloads1mo agoHugging Face23agenticx /DICTranktext1K<n<10K0 likes41 downloads1y agoHugging Face24agenticx /DrugLinkstabular10K<n<100K0 likes37 downloads1y agoHugging Face25Agentic-Payment /agent-spending-incidentsSource of truth and submissions: https://github.com/Pink-Agentic-Payments/agent-spending-incidents (v0.1.0, 2026-10-05). AI Agent Spending Incidents Log A dated, sourced log of publicly reported cases where an AI agent (or LLM-driven automation) spent money it shouldn't have, paid the wrong party, was manipulated into attempting a payment, or leaked payment credentials. Built for researchers, journalists and anyone else who needs concrete, citable examples instead of… See the full description on the dataset page: https://huggingface.co/datasets/Agentic-Payment/agent-spending-incidents.textn<1K0 likes37 downloads5d agoHugging Face26Finance-Agentic-AI /Personal-Finance-Datatabular10K<n<100K0 likes33 downloads10mo agoHugging Face27ethanning /deepresearchgym-agentic-search-logs DeepResearchGym Agentic Search Logs This repository hosts the dataset accompanying the paper “Agentic Search in the Wild” (arXiv: https://arxiv.org/abs/2601.17617). The dataset contains 14M+ search queries collected via DeepResearchGym (DRGym), an open-source search API designed for DeepResearch-style agentic search. For more background on DRGym, see: https://arxiv.org/abs/2505.19253. All records have been anonymized and shuffled to prevent re-identification, and we additionally… See the full description on the dataset page: https://huggingface.co/datasets/ethanning/deepresearchgym-agentic-search-logs.tabulartext-retrieval10M<n<100M1 likes32 downloads8mo agoHugging Face28jamesdborin /Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1-prompt-only Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts.… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Agentic-Indirect-Prompt-Injection-v1-prompt-only.tabular1K<n<10K0 likes32 downloads3mo agoHugging Face29jamesdborin /Nemotron-RL-Agentic-SWE-Pivot-v1-prompt-only Nemotron-RL-Agentic-SWE-Pivot-v1-prompt-only Prompt-only extraction from nvidia/Nemotron-RL-Agentic-SWE-Pivot-v1. Files: prompts.csv: one prompt extraction record per source row. Records include prompt, separated system_prompt, and structured tools when the source row defines available tools. Nested values are JSON-encoded inside CSV cells. summary.md: source row counts, extracted row counts, count deltas, and failed prompt counts. null_or_empty_rows.md: row indexes where… See the full description on the dataset page: https://huggingface.co/datasets/jamesdborin/Nemotron-RL-Agentic-SWE-Pivot-v1-prompt-only.tabular10K<n<100K0 likes31 downloads3mo agoHugging Face30PrismShadow /AgenticRetrievalBench Dataset Card for Agentic Retrieval Benchmark In recent years, there has been a large body of work in the field of text retrieval. However, existing studies are often scattered across different datasets, and their comparisons are partial and fragmented, lacking a comprehensive benchmark for evaluating retrieval performance. This dataset is provided as part of the Agentic Retrieval Benchmark. The project aims to establish a reproducible benchmark for LLM-augmented text retrieval.… See the full description on the dataset page: https://huggingface.co/datasets/PrismShadow/AgenticRetrievalBench.tabular1K<n<10K0 likes29 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.