Team Ai
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01JSALT2026-Conv-AI-Simulator /turnbench-dev-no-backchannel TurnBench Dev - Backchannels Removed A derivative of mundo-ai/turn-benchmark-dev with every majority-annotated backchannel removed from the audio: 1853 backchannels across 38 conversations, 2077.0 seconds in total, cut out of the speaker's own channel and replaced by background noise taken from elsewhere in that same channel. Everything else is the original recording, sample for sample. Same conversations, same duration, same timeline, same annotator tracks, same speech -- only… See the full description on the dataset page: https://huggingface.co/datasets/JSALT2026-Conv-AI-Simulator/turnbench-dev-no-backchannel.audiovoice-activity-detectionn<1K0 likes99 downloads1mo agoHugging Face02PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-flatguard-splittext1K<n<10K0 likes84 downloads2y agoHugging Face03PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite You should mask everything except the last turn. The only part that matters to teach the model is the last turn, as you are teaching it to always output thinking, no matter what the user feeds it. It's setup to be trained like R1: text1K<n<10K1 likes23 downloads2y agoHugging Face04TAUR-dev /D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft This evaluation dataset was created as part of the simple_test__exp_runner_3 experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "TAUR-dev/M-simple_test__exp_runner_3-sft", "tasks": ["countdown_2arg", "countdown_3arg"], "annotators": ["greedy"], "splits": ["test"]… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft.textn<1K0 likes16 downloads1y agoHugging Face05PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-Shuffledtext1K<n<10K0 likes15 downloads6mo agoHugging Face06TAUR-dev /dataset__acronym_generation__simple__4_wordstext1K<n<10K0 likes15 downloads1y agoHugging Face07TAUR-dev /dataset__acronym_generation__simple__5_wordstext1K<n<10K0 likes15 downloads1y agoHugging Face08mlfoundations-dev /dedup_ablation_sim_threshold_10tabular10K<n<100K0 likes12 downloads2y agoHugging Face09TAUR-dev /dataset__acronym_generation__simple__6_wordstext1K<n<10K0 likes12 downloads1y agoHugging Face10mlfoundations-dev /dedup_ablation_sim_threshold_40tabular10K<n<100K0 likes11 downloads2y agoHugging Face11TAUR-dev /dataset__acronym_generation__simple__7_wordstext1K<n<10K0 likes11 downloads1y agoHugging Face12PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-Shuffledtextn<1K0 likes10 downloads2y agoHugging Face13mlfoundations-dev /dedup_ablation_sim_threshold_0tabular10K<n<100K0 likes10 downloads2y agoHugging Face14mlfoundations-dev /dedup_ablation_sim_threshold_nonetabular10K<n<100K0 likes10 downloads2y agoHugging Face15mlfoundations-dev /dedup_ablation_sim_threshold_20tabular10K<n<100K0 likes9 downloads2y agoHugging Face16TAUR-dev /D-simple_test-sft-datatextn<1K0 likes8 downloads1y agoHugging Face17TAUR-dev /D-EVAL__standard_eval_v3__simple_test__exp_runner_2-eval_0 D-EVAL__standard_eval_v3__simple_test__exp_runner_2-eval_0 This evaluation dataset was created as part of the simple_test__exp_runner_2 experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "Qwen/Qwen2.5-1.5b-Instruct", "tasks": ["countdown_2arg", "countdown_3arg"], "annotators": ["greedy"], "splits": ["test"], "dataset_url":… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__simple_test__exp_runner_2-eval_0.textn<1K0 likes7 downloads1y agoHugging Face18TAUR-dev /D-EVAL__standard_eval_v3__back_to_og_mix__simple_mix__rl_eval-eval_rl D-EVAL__standard_eval_v3__back_to_og_mix__simple_mix__rl_eval-eval_rl This evaluation dataset was created as part of the back_to_og_mix__simple_mix__rl_eval experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "TAUR-dev/SIE-back_to_og_mix__simple_retries__sbon-rl", "tasks": ["countdown_2arg", "countdown_3arg", "countdown_4arg"… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__back_to_og_mix__simple_mix__rl_eval-eval_rl.text1K<n<10K0 likes7 downloads1y agoHugging Face19TAUR-dev /D-EVAL__standard_eval_v3__back_to_og_mix__simple_retries__sbon-eval_sft D-EVAL__standard_eval_v3__back_to_og_mix__simple_retries__sbon-eval_sft This evaluation dataset was created as part of the back_to_og_mix__simple_retries__sbon experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "TAUR-dev/M-back_to_og_mix__simple_retries__sbon-sft", "tasks": ["countdown_2arg", "countdown_3arg", "countdown_4arg"… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__back_to_og_mix__simple_retries__sbon-eval_sft.text1K<n<10K0 likes5 downloads1y agoHugging Face20azain /LibriTTS-dev-clean-16khz-mono-loudnorm-100-random-samples-2024-04-18-17-34-39-similaritiestext1K<n<10K0 likes4 downloads2y agoHugging Face21PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all You should mask everything except the last turn. The only part that matters to teach the model is the last turn, as you are teaching it to always output thinking, no matter what the user feeds it. Generation Script This is what I used to generate the dataset, so that it's setup to be trained like R1: import requests import json import time import pandas as pd from datasets import load_dataset from tqdm import… See the full description on the dataset page: https://huggingface.co/datasets/PJMixers-Dev/lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all.textn<1K0 likes4 downloads2y agoHugging Face22PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-flatguard-split-qwqtabularn<1K0 likes4 downloads2y agoHugging Face23TAUR-dev /D-EVAL__simple_eval__cd3arg-sft1ep_grpo_1e6lr_mix_ss_pse_vote_ansrev-sfttext1K<n<10K0 likes4 downloads1y agoHugging Face24TAUR-dev /D-back_to_og_mix__simple_retries__sbon-sft-datatext1K<n<10K0 likes4 downloads1y agoHugging Face25TAUR-dev /D-SFT_C-back_to_og_mix__simple_retries__sbon-sft-datatext1K<n<10K0 likes4 downloads1y agoHugging Face26PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite-filttext1K<n<10K0 likes3 downloads2y agoHugging Face27PJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite-classifiedgatedtext1K<n<10K0 likes1 downloads2y agoHugging Face28PJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-flatguardgatedtext10K<n<100K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.