Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01boomb0om /watermarks-validationimagen<1K3 likes156 downloads4y agoHugging Face02Salesforce /lalm-judge-validation-full-duplex LALM Judge Validation on Full-Duplex Voice Agents Companion dataset for the paper A Reliability Assessment of LALM Audio Judges for Full-Duplex Voice Agents. This repository contains the anonymised ratings, adversarial-defect recall tables, JSON schemas, and analysis scripts used to produce every headline number, table, and figure in that paper. Summary 209 rated stereo sessions: 152 full-duplex agent-client conversations across 13 accent-and-condition strata… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/lalm-judge-validation-full-duplex.tabularaudio-classification1K<n<10K2 likes93 downloads3mo agoHugging Face03vaskers5 /latent_upscale_validationimagen<1K0 likes57 downloads11mo agoHugging Face04sewlore /synthetic-fabric-csv-validation-cases Sewlore Synthetic Fabric CSV Validation Cases All 72 records are invented software inputs. This is a small, deterministic CSV-validation teaching corpus. It contains no physical fabric measurements, garment trials, personal data, photographs or washing observations. No model was trained or evaluated. AI assistance was used to prepare cases and documentation; labels were captured by executing a frozen public Python package. Use it to learn a specific parser contract, compare… See the full description on the dataset page: https://huggingface.co/datasets/sewlore/synthetic-fabric-csv-validation-cases.textn<1K0 likes49 downloads16h agoHugging Face05Noddybear /fisher-validation-resultsimagen<1K0 likes26 downloads5mo agoHugging Face06STEVENZHANG904 /Build_Bench_Validation_DataThis repository contains the validation set of BuildBench paper. It contains 70 data samples. textn<1K0 likes18 downloads1y agoHugging Face07MihaiIonascu /dreadit-validationtextn<1K0 likes17 downloads4y agoHugging Face08beddi /dataset-train-validationtext10K<n<100K0 likes17 downloads2y agoHugging Face09MihaiIonascu /Azure_IaC_validationtextn<1K0 likes15 downloads3y agoHugging Face10Chenyang036 /Validation_Setstext10K<n<100K0 likes15 downloads1y agoHugging Face11samim2024 /medical-chatbot-validation-datasettextn<1K0 likes12 downloads1y agoHugging Face12AITeamUIT /eval-gliner2-ner-bc5cdr-boundary-smoothing-validationtabularn<1K0 likes12 downloads3mo agoHugging Face13Javtor /biomedical-topic-categorization-validationtext100K<n<1M1 likes10 downloads4y agoHugging Face14AiDoc /CT-MRI_metrics_validation_datasettext10K<n<100K1 likes10 downloads4y agoHugging Face15loyoladatamining /usajobs_validation USAJOBS Dataset (Validation Sample) Dataset Description The USAJOBS Dataset is a comprehensive collection of federal job postings from January 2017 through March 2026. This dataset includes full-text job descriptions, and structured metadata (job title and employer). This particular dataset presents a sample of sentence-level data from the corpus, tagged with task, skill, and AI attributes. Dataset Structure The dataset contains 20k sentences… See the full description on the dataset page: https://huggingface.co/datasets/loyoladatamining/usajobs_validation.tabular10K<n<100K0 likes10 downloads4mo agoHugging Face16Abdulrahman44 /Validationtabular10M<n<100M0 likes9 downloads3y agoHugging Face17Mira-Network /ensemble-validation Ensemble Evaluation Data Dataset Summary The Learnrite Evaluation Data is a comprehensive question bank designed for evaluating AI models on complex, real-world questions derived from India’s Civil Services examination — widely regarded as one of the toughest competitive exams globally. The dataset features multiple-choice questions (MCQs) covering topics such as the Indian Constitution, governance, and administrative functions. This makes it a particularly challenging… See the full description on the dataset page: https://huggingface.co/datasets/Mira-Network/ensemble-validation.textn<1K3 likes9 downloads2y agoHugging Face18nqzfaizal77ai /squad-qa-validationThis dataset processed version of the SQuAD dataset, which is provided by Hugging Face. The SQuAD dataset is a collection of questions and answers derived from a set of Wikipedia articles, designed for machine reading comprehension tasks. Dataset source: https://huggingface.co/datasets/rajpurkar/squad text10K<n<100K0 likes9 downloads2y agoHugging Face19AITeamUIT /eval-gliner2-ner-ncbi_disease-boundary-smoothing-validationtabularn<1K0 likes9 downloads3mo agoHugging Face20AITeamUIT /eval-gliner2-ner-mit_restaurant-affine-boundary-smoothing-validationtabularn<1K0 likes9 downloads3mo agoHugging Face21JuniorBueno /doctor_validationtext10K<n<100K0 likes8 downloads2y agoHugging Face22AITeamUIT /eval-gliner2-ner-wnut2017-affine-boundary-smoothing-validationtabularn<1K0 likes7 downloads3mo agoHugging Face23AITeamUIT /eval-gliner2-ner-ncbi_disease-affine-boundary-smoothing-validationtabularn<1K0 likes7 downloads3mo agoHugging Face24hojzas /setfit-proj8-multilabel_2_validationtabularn<1K0 likes6 downloads3y agoHugging Face25AITeamUIT /eval-gliner2-ner-ontonotes5-boundary-smoothing-validationtabularn<1K0 likes6 downloads3mo agoHugging Face26AITeamUIT /eval-gliner2-ner-conll2003-affine-validationtabularn<1K0 likes6 downloads3mo agoHugging Face27jacyanthis /CompanionSim-Validation CompanionSim-Validation 70 real-world conversations annotated by two groups: 168 annotators in the US (CompanionSim-Validation-US.csv) and 998 annotators from the US, UK, India, and Nigeria (CompanionSim-Validation-Multi.csv). tabular1K<n<10K0 likes6 downloads2mo agoHugging Face28Slepp /validationvalidation set textn<1K0 likes5 downloads4y agoHugging Face29zluvolyote /Dream_NLP_Validationtabular100K<n<1M0 likes5 downloads4y agoHugging Face30hojzas /proj8-label-validationtextn<1K0 likes5 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.