Team Ai
20 results

dev-sim

JSALT2026-Conv-AI-Simulator /turnbench-dev-no-backchannel TurnBench Dev - Backchannels Removed A derivative of mundo-ai/turn-benchmark-dev with every majority-annotated backchannel removed from the audio: 1853 backchannels across 38 conversations, 2077.0 seconds in total, cut out of the speaker's own channel and replaced by background noise taken from elsewhere in that same channel. Everything else is the original recording, sample for sample. Same conversations, same duration, same timeline, same annotator tracks, same speech -- only… See the full description on the dataset page: https://huggingface.co/datasets/JSALT2026-Conv-AI-Simulator/turnbench-dev-no-backchannel.audiovoice-activity-detectionn<1K0 likes160 downloads28d agoHugging FacePJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT-flatguard-splittext1K<n<10K0 likes84 downloads2y agoHugging FaceRaymond-dev-546730 /Simple-MathSteps-90K Introducing Simple-MathSteps-90K: An open source dataset of 93,325 elementary math problems with step-by-step solutions and multiple choice answers. Designed to enhance mathematical reasoning in models ranging from 1B to 13B parameters. Key Features 93,325 Math Problems: Generated by paraphrasing the AQuA-RAT dataset using Qwen3 4B Instruct 2507, with a focus on consistency and quality. Detailed Step-by-Step Solutions: Clear reasoning that breaks down problems… See the full description on the dataset page: https://huggingface.co/datasets/Raymond-dev-546730/Simple-MathSteps-90K.3 likes76 downloads2mo agoHugging FacePJMixers-Dev /lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite lemonilia_LimaRP-Only-NonSus-Simple-CustomShareGPT-qwq-all-aphrodite You should mask everything except the last turn. The only part that matters to teach the model is the last turn, as you are teaching it to always output thinking, no matter what the user feeds it. It's setup to be trained like R1: text1K<n<10K1 likes25 downloads2y agoHugging FacePJMixers-Dev /lemonilia_LimaRP-Simple-CustomShareGPT1K<n<10K0 likes20 downloads6mo agoHugging FaceTAUR-dev /D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft This evaluation dataset was created as part of the simple_test__exp_runner_3 experiment using the SkillFactory experiment management system. Experiment Tracking 🔗 View complete experiment details: Experiment Tracker Dataset Evaluation Details {"model": "TAUR-dev/M-simple_test__exp_runner_3-sft", "tasks": ["countdown_2arg", "countdown_3arg"], "annotators": ["greedy"], "splits": ["test"]… See the full description on the dataset page: https://huggingface.co/datasets/TAUR-dev/D-EVAL__standard_eval_v3__simple_test__exp_runner_3-eval_sft.textn<1K0 likes18 downloads1y agoHugging Face