Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01aisingapore /NLU-Sentiment-Analysisgated SEA Sentiment Analysis SEA Sentiment Analysis evaluates a model's ability to identify the sentiment polarity of a text. It is sampled from NusaX for Indonesian, Javanese, and Sundanese, IndicSentiment for Tamil, Wisesight Sentiment for Thai, and UIT-VSFC for Vietnamese. Supported Tasks and Leaderboards SEA Sentiment Analysis is designed for evaluating chat or instruction-tuned large language models (LLMs). It is part of the SEA-HELM leaderboard from AI Singapore.… See the full description on the dataset page: https://huggingface.co/datasets/aisingapore/NLU-Sentiment-Analysis.texttext-generation1K<n<10K0 likes1.1k downloads10mo agoHugging Face02daaain /swebench-verified-deepseek-v4-flash-failure-analysis SWE-bench Verified runs & failure analysis — DeepSeek-V4-flash (local) × mini-swe-agent Per-instance analysis of SWE-bench Verified runs of a locally-served DeepSeek-V4-flash model driven by mini-swe-agent, graded with the official SWE-bench harness. Each instance carries the full agent trajectory, a readable transcript, the submitted patch, the harness test output, deterministic metrics, and a hand-verified qualitative root-cause diagnosis. Current numbers (resolve rates… See the full description on the dataset page: https://huggingface.co/datasets/daaain/swebench-verified-deepseek-v4-flash-failure-analysis.tabulartext-generationn<1K0 likes306 downloads4mo agoHugging Face03jakeatx /qwen36-mtp-turbo-kv-analysis Qwen3.6 MTP Turbo KV Runtime Analysis This repository is a curated analysis artifact for local Qwen3.6-35B-A3B MTP GGUF inference experiments on Windows CUDA. It compares clean MTP llama.cpp, QuinsZouls llama-next TurboQuant, and the completed subset of Atomic TurboQuant runs under a fixed 64k context, MoE CPU offload, and Unsloth-aligned sampling settings. The raw benchmark runs included incomplete and capability-incompatible rows. This repo keeps only completed, comparable… See the full description on the dataset page: https://huggingface.co/datasets/jakeatx/qwen36-mtp-turbo-kv-analysis.imagetext-generationn<1K1 likes263 downloads5mo agoHugging Face04danielrosehill /Open-Router-API-Pricing-Analysis OpenRouter API Pricing Analysis Dataset Overview This dataset provides a point-in-time capture of pricing and parameters for LLMs available through the OpenRouter API for inference. Contents Raw Data (raw/) Contains the original data extracted from the OpenRouter API, including: Model pricing (input/output token costs) Model parameters and specifications Computed fields such as output/input token price ratios Enhanced Data (hf-enhanced/)… See the full description on the dataset page: https://huggingface.co/datasets/danielrosehill/Open-Router-API-Pricing-Analysis.texttext-generation1K<n<10K0 likes209 downloads11mo agoHugging Face05lylybig8 /routing_analysis-finetuning-data routing_analysis finetuning data A mirror of routing_analysis/finetuning/data/: the JSONL training splits used in the multilingual MoE language-expansion experiments, plus the generator scripts, runners, manifests and logs that produced them. Layout Path Contents Size splits/ Document-count tiers {lang}_{tier}.jsonl, {lang}_manifest.json, splits_summary.json ~36 GB token_budget_splits/ Token-budget tiers {lang}_{1Btok,1p5Btok,2Btok}.jsonl and… See the full description on the dataset page: https://huggingface.co/datasets/lylybig8/routing_analysis-finetuning-data.text-generation10M<n<100M0 likes195 downloads14d agoHugging Face06rodriguescarson /adaption-market-analysis-invented-clean-1000 Market Analysis Q&A (Invented) Market-analysis questions with long-form answers. Rows 1,000 Domain market analysis Format data.parquet, one row per example Licence other Built for supervised fine-tuning (SFT) experiments on Adaption AutoScientist Columns Column Description original_prompt The prompt (user turn) as uploaded. original_completion The target response as uploaded. enhanced_prompt Empty in this dataset.… See the full description on the dataset page: https://huggingface.co/datasets/rodriguescarson/adaption-market-analysis-invented-clean-1000.tabulartext-generation1K<n<10K0 likes66 downloads14d agoHugging Face07rodriguescarson /adaption-si-dimensional-analysis-12k SI Dimensional Analysis Dimensional reasoning in SI base units: is a relation dimensionally consistent, reduce a combination to base units, find the term that cannot belong in a sum, and decide whether two combinations share dimensions. Rows 12,000 Domain physics Format data.parquet, one row per example Licence cc-by-4.0 Built for supervised fine-tuning (SFT) experiments on Adaption AutoScientist Columns Column Description… See the full description on the dataset page: https://huggingface.co/datasets/rodriguescarson/adaption-si-dimensional-analysis-12k.tabulartext-generation10K<n<100K0 likes63 downloads14d agoHugging Face08yatin-superintelligence /Adversarial-Agent-Intent-Safety-Analysis-240Kgated Adversarial Agent Intent Safety Analysis 240K Abstract The Adversarial-Agent-Intent-Safety-Analysis-240K is a deterministically structured dataset featuring 242,454 context-rich adversarial prompts and safety evaluations. Engineered strictly for training frontier command-and-control models, guardrail classifiers, and red-teaming agents, it encourages models to parse multi-layered intention across 126 critical risk vectors. This design trains models to decouple the surface… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/Adversarial-Agent-Intent-Safety-Analysis-240K.texttext-classification100K<n<1M13 likes55 downloads7mo agoHugging Face09sfd-anonymous /sefd-archive-100k-analysis-sample-qwen3-20260524 SEFD Archive 100k Analysis Sample Qwen3 20260524 Retained artifacts for the completed archive-wide 100,000-filing Stanford EDGAR Filings Dataset (SEFD) analysis sample used in the arXiv paper update. The sample contains 2,971,490,909 final SEFD tokens, counted with the Qwen3-1.7B tokenizer. This repository is a new versioned artifact and intentionally does not replace the earlier sfd-archive-100k-analysis-sample repository used for the original conference submission. Included:… See the full description on the dataset page: https://huggingface.co/datasets/sfd-anonymous/sefd-archive-100k-analysis-sample-qwen3-20260524.text-generation0 likes52 downloads5mo agoHugging Face10rodriguescarson /adaption-market-analysis-invented-1k Market Analysis Q&A (Invented) Market-analysis questions with long-form answers. Rows 1,000 Domain market analysis Format data.parquet, one row per example Licence other Built for supervised fine-tuning (SFT) experiments on Adaption AutoScientist Columns Column Description enhanced_prompt Prompt after Adaption processing (rewrite or augmentation). enhanced_completion Response after Adaption processing (rewrite or… See the full description on the dataset page: https://huggingface.co/datasets/rodriguescarson/adaption-market-analysis-invented-1k.texttext-generation1K<n<10K0 likes52 downloads14d agoHugging Face11miscusi /adaption-market-analysis-sec Market Analysis & News Instruction Dataset (SEC XBRL-grounded) Instruction-tuning data for financial analysis — fundamentals, growth and ratio arithmetic, trend and risk reading, filing navigation and comparability caveats — built from real XBRL facts, with every stated figure independently re-derived. Built for the Adaption Labs AutoScientist Challenge Part 2, Market Analysis & News track. What is in it Rows 5,068 (4,501 train / 567 eval) Task… See the full description on the dataset page: https://huggingface.co/datasets/miscusi/adaption-market-analysis-sec.texttext-generation1K<n<10K0 likes50 downloads2mo agoHugging Face12yothinS /Customer_Behavior_Analysis Customer Behavior Analysis & CRM Intelligence Dataset (Thailand Context) A specialized instruction-tuning dataset designed to train Large Language Models (LLMs) to serve as an Internal CRM Intelligence Copilot tailored specifically for the Thai market and consumer landscape. The dataset bridges quantitative transaction logs (RFM, usage telemetry) with qualitative customer psychological theories within the local Thai business ecosystem (e.g., LINE OA interactions… See the full description on the dataset page: https://huggingface.co/datasets/yothinS/Customer_Behavior_Analysis.texttext-generation1K<n<10K0 likes47 downloads6d agoHugging Face13stindardlogic /financial-analysis-sft-100k Financial Analysis SFT (100K) 100,000 ShareGPT conversations demonstrating expert-level financial analysis across DCF modeling, unit economics, LBO analysis, credit analysis, earnings interpretation, comparable company analysis, and financial ratio analysis. Motivation Financial AI is a critical enterprise capability — investment analysts, CFOs, startup founders, and finance teams need models that can reason through complex financial questions with the precision… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/financial-analysis-sft-100k.texttext-generation100K<n<1M1 likes41 downloads3mo agoHugging Face14EngineeringWays /Circuit-Analysis-Reasoning-Sample ⚡ EngineeringWays Data Lab: Circuit Analysis Reasoning Dataset (Free Sample) This is a free 50-item sample of the EngineeringWays Circuit Analysis Reasoning Dataset. It is designed specifically for fine-tuning Large Language Models (LLMs) in advanced STEM problem-solving, featuring strict Chain-of-Thought (CoT) reasoning. Want the complete, deduplicated 592-item master dataset? 👉 Get the LoRA-Ready Master File on Payhip 🚀 Dataset Overview Most math and physics… See the full description on the dataset page: https://huggingface.co/datasets/EngineeringWays/Circuit-Analysis-Reasoning-Sample.texttext-generationn<1K1 likes39 downloads6mo agoHugging Face15chn123 /bimcv-analysisgated BIMCV-R 500-Series Multi-Model Analysis & Explainability Dataset This repository contains end-to-end multi-model AI evaluation, lung cancer risk modeling, explainability maps (Grad-CAM), and automated radiology report generation for 500 randomly sampled physician-labeled chest CT series from the cyd0806/BIMCV-R dataset. The cohort was evaluated across three state-of-the-art chest CT deep learning models: Sybil (5-Seed Ensemble): 1-to-6 year lung cancer risk probability… See the full description on the dataset page: https://huggingface.co/datasets/chn123/bimcv-analysis.texttabular-classification10K<n<100K0 likes37 downloads1d agoHugging Face16sh111111111111111 /cve-analysis CVE & Vulnerability Analysis Dataset A comprehensive vulnerability analysis and CVE research dataset. Each row is a detailed security analysis covering root cause, exploitation methodology, detection rules (Sigma/Splunk/Suricata), CVSS v3.1 scoring, MITRE ATT&CK mapping, and remediation guidance — verified by the same model in an independent review pass. Overview This dataset contains 9,999 structured vulnerability analyses across 20 security domains. Unlike simple… See the full description on the dataset page: https://huggingface.co/datasets/sh111111111111111/cve-analysis.texttext-generation1K<n<10K1 likes36 downloads7mo agoHugging Face17sfd-anonymous /sfd-archive-100k-analysis-sample SFD Archive 100k Analysis Sample Retained artifacts for the completed archive-wide 100k-filing SFD analysis sample. The run processed 99,895 filings, produced 99,607 successful parses, and contains 2,698,335,638 final SFD tokens. This repository contains aggregate run metadata, processed accessions, and paper-analysis CSV/JSON metrics. Per-filing markdown outputs and temporary raw SEC downloads are not included. text-generation0 likes31 downloads5mo agoHugging Face18Papajams /autoscientist-market-analysis-lenitnes-dataset autoscientist-market-analysis-lenitnes-dataset The adapted dataset used to fine-tune Papajams/autoscientist-market-analysis-lenitnes for the Adaption Labs AutoScientist Challenge Part 2 (Market-Analysis & News category). Composition Total rows 27,965 Real seed rows (production DB) 1002 Unique source signals 272 Augmented rows (~19K domain + ~8K diversity) AutoScientist-augmented Seed provenance (real data): the lenitnes production platform… See the full description on the dataset page: https://huggingface.co/datasets/Papajams/autoscientist-market-analysis-lenitnes-dataset.texttext-generation10K<n<100K0 likes27 downloads2mo agoHugging Face19316usman /sentiment-analysis SENTIMENT_ANALYSIS A preference dataset for SENTIMENT_ANALYSIS, harvested from real, human-labelled sources and curated by an automated harvesting harness with an LLM quality gate. Format Standard preference / DPO schema — each row: column meaning prompt the request (originally text) chosen the human-preferred response rejected a worse response to the same prompt source the dataset/URL the row was harvested from Splits 80/10/10… See the full description on the dataset page: https://huggingface.co/datasets/316usman/sentiment-analysis.texttext-generation1K<n<10K0 likes23 downloads2d agoHugging Face20AhmedBou /NCSS_2023_Data_Analysistexttoken-classificationn<1K0 likes20 downloads3y agoHugging Face21mangesh-ux /logistics-cx-transcript-analysis-chatml OmniCX Logistics CX Dataset (Research Preview) Dataset Summary This dataset is designed for structured extraction of logistics and customer-experience (CX) signals from multi-turn support conversations. Each record uses ChatML-style messages with: a fixed system instruction a user transcript an assistant JSON payload matching LogisticsCXMetrics This release is a research preview and should not be treated as a production-certified benchmark. Project repository:… See the full description on the dataset page: https://huggingface.co/datasets/mangesh-ux/logistics-cx-transcript-analysis-chatml.texttext-generationn<1K0 likes19 downloads7mo agoHugging Face22twistshan /realistic-niah-count-mechanism-analysis Realistic NIAH count mechanism analysis Version 2 stores the paired geometry panel once. The default geometry_shared configuration contains 300 unique V4.4 stimulus rows: 200 discovery rows (seeds 1234-1253) and 100 held-out confirmation rows (seeds 1254-1263), with counts 1-10 balanced within every seed. Each pair_id is now one row rather than two duplicated mode rows. The common row contains the passage, gold records, slots, active needle spans, hard negatives, design metadata… See the full description on the dataset page: https://huggingface.co/datasets/twistshan/realistic-niah-count-mechanism-analysis.tabulartext-generationn<1K0 likes19 downloads2mo agoHugging Face23haesleinhuepf /bio-image-analysis-qa Dataset Card for bio-image-analysis-qa This dataset contains questions and answers for analysing biological microscopy imaging data using python. Dataset Details Dataset Description Questions and answers provided in this repository are centered around the topic, how to process imaging data using Python. Curated by: Robert Haase License: CC-BY 4.0 Dataset Sources and Processing This dataset was derived from the Bio-image Analysis Notebooks which… See the full description on the dataset page: https://huggingface.co/datasets/haesleinhuepf/bio-image-analysis-qa.texttext-generationn<1K1 likes18 downloads2y agoHugging Face24massines3a /chainscope-analysis ChainScope Qwen3-8B Faithfulness Analysis Dataset This dataset contains Chain-of-Thought (CoT) faithfulness evaluation data for Qwen3-8B, including hidden state activations, labeled sentences, and evaluation results. Dataset Description We evaluated CoT faithfulness using the ChainScope methodology: Generate CoT responses for comparison questions (e.g., "Is A > B?") Generate "reversed" CoT responses for the opposite question ("Is B > A?") Compare whether the model's… See the full description on the dataset page: https://huggingface.co/datasets/massines3a/chainscope-analysis.texttext-generation10K<n<100K0 likes18 downloads8mo agoHugging Face25MK4-Research /VAB-vulnerability-analysis-benchmark FBE and VAB Two small benchmarks for security code analysis. Both grade without an LLM judge, so runs are cheap and repeatable. FBE (find-the-bug) 14 code snippets, each with one planted vulnerability. Ask the model to analyze the code, then check whether it actually found the flaw. Grading uses concept groups: the answer has to contain at least one synonym from every required group. Four numbers come out: found, did it identify the real vulnerability (this is… See the full description on the dataset page: https://huggingface.co/datasets/MK4-Research/VAB-vulnerability-analysis-benchmark.textquestion-answeringn<1K0 likes18 downloads2mo agoHugging Face26kilicai /turkish-doc-summary-review-analysis-30k-jsonl ⚠️ Superseded by v2 Bu v1 dataset'te exact duplicate yoktu; ancak belge ve cevap şablonları fazla tekrar ediyordu. Güncel v2 sürümünü kullanın: https://huggingface.co/datasets/kilicai/turkish-doc-summary-review-analysis-30k-jsonl-v2 Generated by ML Intern This dataset repository was generated by ML Intern, an agent for machine learning research and development on the Hugging Face Hub. Try ML Intern: https://smolagents-ml-intern.hf.space Source code:… See the full description on the dataset page: https://huggingface.co/datasets/kilicai/turkish-doc-summary-review-analysis-30k-jsonl.texttext-generation10K<n<100K0 likes17 downloads5mo agoHugging Face27Odelolasolomon /mistakes-analysis Nanbeige-4-3B-Base: Semantic Blind Spots & Mistake Analysis Overview This dataset is a curated collection of 10+ high-fidelity semantic failures identified during the evaluation of the Nanbeige/Nanbeige4-3B-Base model. While the Nanbeige model is highly capable for its size, my testing revealed specific "blind spots" in logical reasoning, strict constraint adherence, and mathematical precision. This dataset serves as a benchmark for where the model currently fails… See the full description on the dataset page: https://huggingface.co/datasets/Odelolasolomon/mistakes-analysis.texttext-generationn<1K0 likes16 downloads8mo agoHugging Face28VaisakhKrishna /Emotional_Sentiment_AnalysisEmotional Sentiment Analysis Dataset for LLaMA-2 Fine-tuning (The formatted version can be directly used for fine tuning which contain only the formatted text, while the dataset.csv contain all the text, emotion, response and the formatted text) This dataset contains conversational data for training and fine-tuning language models for emotional sentiment analysis and response generation. The dataset includes user inputs, their corresponding emotional states, and tailored chatbot responses… See the full description on the dataset page: https://huggingface.co/datasets/VaisakhKrishna/Emotional_Sentiment_Analysis.texttext-classification1K<n<10K2 likes15 downloads2y agoHugging Face29kilicai /turkish-doc-summary-review-analysis-30k-jsonl-v2 Turkish Document Summary Review Analysis 30K JSONL v2 Tek dosya: train.jsonl. Bu v2 sürümü, v1'de görülen tekrar sorununu çözmek için yeniden üretildi: Daha fazla belge türü: proje önerisi, tutanak, denetim notu, şikâyet dosyası, politika taslağı, saha raporu, bütçe değerlendirmesi, risk kayıt formu, karar destek belgesi, olay inceleme raporu vb. Daha fazla alt görev: 30 farklı task_type. Exact duplicate + semantic template duplicate kontrolü. Cevap şablonları belgeye özel risk… See the full description on the dataset page: https://huggingface.co/datasets/kilicai/turkish-doc-summary-review-analysis-30k-jsonl-v2.texttext-generation10K<n<100K0 likes15 downloads5mo agoHugging Face30TatarNLPWorld /tatar-news-analysis-multiclassgated Dataset Card for Tatar News Multiclass Classification Dataset Details Dataset Description The Tatar News Multiclass Classification Dataset contains 86,963 Tatar language news articles classified into 9 distinct topic categories. Each entry includes the full article content, title, category (numeric label and text label), source URL, publication date, and content length. The dataset is specifically designed for training and evaluating multi-class… See the full description on the dataset page: https://huggingface.co/datasets/TatarNLPWorld/tatar-news-analysis-multiclass.tabulartext-classification10K<n<100K0 likes14 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.