Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01mahiatlinux /Reflection-Dataset-ShareGPT-v2 Simple "Reflection" method dataset inspired by mattshumer This is the ShareGPT version. Find prompt and response pair dataset here This dataset was synthetically generated using Glaive AI. There have been structure improvements and added more rows. text1K<n<10K15 likes314 downloads2y agoHugging Face02bigai-nlco /ReflectionEvoGithub Repo for ReflectEvo: https://github.com/bigai-nlco/ReflectEvo Arxiv Paper for ReflectEvo: https://arxiv.org/abs/2505.16475 textquestion-answering100K<n<1M12 likes246 downloads1y agoHugging Face03SenseLLM /ReflectionSeq-DS ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation 📄 Paper • 🏠 Repo • 🤖 Models • 📚 Datasets Introduction ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details! Models Model Checkpoint Size HumanEval (+) MBPP (+)… See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-DS.texttext-generation10K<n<100K5 likes93 downloads2y agoHugging Face04mahiatlinux /Reflection-Dataset-v2 Second version of a simple "Reflection" method dataset inspired by mattshumer This is the prompt and response version. Find ShareGPT version here This dataset was synthetically generated using Glaive AI. There have been structure improvements and added more rows. text1K<n<10K37 likes63 downloads2y agoHugging Face05SenseLLM /ReflectionSeq-GPT ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation 📄 Paper • 🏠 Repo • 🤖 Models • 📚 Datasets Introduction ReflectionCoder is a novel approach that effectively leverages reflection sequences constructed by integrating compiler feedback to improve one-off code generation performance. Please refer to our paper and repo for more details! Models Model Checkpoint Size HumanEval (+) MBPP (+)… See the full description on the dataset page: https://huggingface.co/datasets/SenseLLM/ReflectionSeq-GPT.texttext-generation10K<n<100K5 likes61 downloads2y agoHugging Face06dougalldeepmind /2026-08-04-qwen36-self-reflection-20-80-train ⚠️ SUPERSEDED — do not train from this bundle Built 2026-08-04 under the old rendering policy, where Qwen3.6 emitted a <think> block on the final assistant turn only. The repository has since moved to preserve-thinking rendering, in which every assistant turn carries a think block (real trace, or the empty marker). Both files here are stale as a result: mixture.jsonl — rendered under the old policy, so it trains different strings than the current pipeline produces. It also… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-04-qwen36-self-reflection-20-80-train.text1K<n<10K0 likes54 downloads2mo agoHugging Face07open-llm-leaderboard /SenseLLM__ReflectionCoder-CL-34B-detailsgated Dataset Card for Evaluation run of SenseLLM/ReflectionCoder-CL-34B Dataset automatically created during the evaluation run of model SenseLLM/ReflectionCoder-CL-34B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/SenseLLM__ReflectionCoder-CL-34B-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face08open-llm-leaderboard /EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Alpaca-Llama3.1.08-8B-C-R1-KTO-Reflection-details.tabular10K<n<100K0 likes37 downloads2y agoHugging Face09mahiatlinux /Reflection-Dataset-v1 V2 is out!!! V2 Simple "Reflection" method dataset inspired by mattshumer This is the prompt and response version. Find ShareGPT version here This dataset was synthetically generated using Glaive AI. text1K<n<10K20 likes36 downloads2y agoHugging Face10open-llm-leaderboard /mattshumer__Reflection-Llama-3.1-70B-detailsgated Dataset Card for Evaluation run of mattshumer/Reflection-Llama-3.1-70B Dataset automatically created during the evaluation run of model mattshumer/Reflection-Llama-3.1-70B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mattshumer__Reflection-Llama-3.1-70B-details.tabular10K<n<100K0 likes34 downloads2y agoHugging Face11Aculi /Reflection-AlpacaA mix of multiple other Datasets and also some Selfmade and GPT-4o Prompts. text10K<n<100K1 likes34 downloads2y agoHugging Face12open-llm-leaderboard /olabs-ai__reflection_model-detailsgated Dataset Card for Evaluation run of olabs-ai/reflection_model Dataset automatically created during the evaluation run of model olabs-ai/reflection_model The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/olabs-ai__reflection_model-details.tabular10K<n<100K0 likes33 downloads2y agoHugging Face13open-llm-leaderboard /EpistemeAI2__Fireball-Llama-3.1-8B-Philos-Reflection-detailsgated Dataset Card for Evaluation run of EpistemeAI2/Fireball-Llama-3.1-8B-Philos-Reflection Dataset automatically created during the evaluation run of model EpistemeAI2/Fireball-Llama-3.1-8B-Philos-Reflection The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/EpistemeAI2__Fireball-Llama-3.1-8B-Philos-Reflection-details.tabular10K<n<100K0 likes33 downloads2y agoHugging Face14open-llm-leaderboard /SenseLLM__ReflectionCoder-DS-33B-detailsgated Dataset Card for Evaluation run of SenseLLM/ReflectionCoder-DS-33B Dataset automatically created during the evaluation run of model SenseLLM/ReflectionCoder-DS-33B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/SenseLLM__ReflectionCoder-DS-33B-details.tabular10K<n<100K0 likes31 downloads2y agoHugging Face15appvoid /reflection-v1the dataset was generated using synthetic data from llama-3.1-70b following the format of the popular reflection model to improve reasoning on small language models to ensure diversity, the model can be teached to learn when to use shorter or longer text to reflect on the actual task column '0' offers shorter responses while column '1' offer longer ones textn<1K0 likes28 downloads2y agoHugging Face16ericflo /unnaturalhermes-reflections-100ktext100K<n<1M2 likes26 downloads3y agoHugging Face17jiazhengli /DARS_synthethsis_reflection DARS: Dual-Model Verbal Reflection Datasets This repository contains the training datasets for the DARS (Dual-model Reflective Scoring) framework, a novel approach for automated student answer scoring that uses verbal reflection at inference time. Overview The DARS framework employs two specialized models working in tandem: Reasoner: Generates initial assessments and refines them based on feedback Critic: Provides targeted verbal reflections and determines when reasoning… See the full description on the dataset page: https://huggingface.co/datasets/jiazhengli/DARS_synthethsis_reflection.texttext-generation10K<n<100K0 likes26 downloads1y agoHugging Face18mahiatlinux /Reflection-Dataset-ShareGPT-v1 V2 is out!!! V2 Simple "Reflection" method dataset inspired by mattshumer This is the ShareGPT version. Find prompt and response pair dataset here This dataset was synthetically generated using Glaive AI. text1K<n<10K8 likes23 downloads2y agoHugging Face19PJMixers-Dev /Weyaxi_HelpSteer-filtered-Reflection-Gemini-1.5-Flash-ShareGPTSystem prompt taken from here and slightly modified. 4 examples were generated with GPT-4o and then slightly modified. The examples were sent to Gemini-1.5-Flash followed by the real user turn. Any samples which had responses which were stopped with finish_reason: "SAFETY" were skipped, and responses which did not contain one of each start/stop tag were regenerated. If it failed to generate within 3 tries the sample was skipped. model = genai.GenerativeModel( "models/gemini-1.5-flash"… See the full description on the dataset page: https://huggingface.co/datasets/PJMixers-Dev/Weyaxi_HelpSteer-filtered-Reflection-Gemini-1.5-Flash-ShareGPT.text1K<n<10K3 likes23 downloads2y agoHugging Face20march228 /grok-reflection-cot-ru march228/grok-reflection-cot-ru Russian synthetic dataset with question, internal thought text, and final answer. What is inside Rows: 4190 Split: train Main fields: question thought_text answer thought1..thought5 model task_type reflection_count Format The dataset is stored as train.jsonl. thought_text is the joined internal monologue with blank lines between thought blocks.thought1..thought5 preserve the original segmented form from the SQLite source.… See the full description on the dataset page: https://huggingface.co/datasets/march228/grok-reflection-cot-ru.tabulartext-generation1K<n<10K1 likes22 downloads7mo agoHugging Face21galaxychen /Overthinking-FCS-Reflectiontext1K<n<10K1 likes19 downloads10mo agoHugging Face22open-llm-leaderboard /glaiveai__Reflection-Llama-3.1-70B-detailsgated Dataset Card for Evaluation run of glaiveai/Reflection-Llama-3.1-70B Dataset automatically created during the evaluation run of model glaiveai/Reflection-Llama-3.1-70B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/glaiveai__Reflection-Llama-3.1-70B-details.tabular10K<n<100K0 likes18 downloads2y agoHugging Face23open-llm-leaderboard /mosama__Qwen2.5-1.5B-Instruct-CoT-Reflection-detailsgated Dataset Card for Evaluation run of mosama/Qwen2.5-1.5B-Instruct-CoT-Reflection Dataset automatically created during the evaluation run of model mosama/Qwen2.5-1.5B-Instruct-CoT-Reflection The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mosama__Qwen2.5-1.5B-Instruct-CoT-Reflection-details.tabular10K<n<100K0 likes18 downloads2y agoHugging Face24open-llm-leaderboard /Xiaojian9992024__Reflection-L3.2-JametMiniMix-3B-detailsgated Dataset Card for Evaluation run of Xiaojian9992024/Reflection-L3.2-JametMiniMix-3B Dataset automatically created during the evaluation run of model Xiaojian9992024/Reflection-L3.2-JametMiniMix-3B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Xiaojian9992024__Reflection-L3.2-JametMiniMix-3B-details.tabular10K<n<100K0 likes16 downloads2y agoHugging Face25smjain /api_graph_reflectiontabularn<1K0 likes15 downloads2y agoHugging Face26stvlynn /Reflection-Chinese-Dataset Reflection-Chinese-Dataset·Reflection中文数据集 Based on mahiatlinux/Reflection-Dataset-v2, translated using RA Translation Tool text1K<n<10K10 likes14 downloads2y agoHugging Face27isaiahbjork /reflection-spelling-puzzles-sharegptgatedtext1K<n<10K1 likes14 downloads2y agoHugging Face28achiepatricia /han-autonomous-decision-reflection-dataset-v1 Autonomous Decision Reflection Dataset This dataset records autonomous decisions made by humanoid AI and the reflective reasoning behind those decisions. It supports transparency and explainability in autonomous systems. Use Cases Explainable AI Decision auditing Safety analysis Fields decision_context chosen_action reasoning_summary outcome Part of Humanoid Network (HAN) License MIT textn<1K0 likes14 downloads8mo agoHugging Face295CD-AI /Vietnamese-mahiatlinux-Reflection-Dataset-ShareGPT-v2-gg-translatedtextquestion-answering1K<n<10K3 likes13 downloads2y agoHugging Face30isaiahbjork /reflection-40k-sharegptgated Datasets mahiatlinux/Reflection-Dataset-ShareGPT-v2: 9171 isaiahbjork/reflection-scienceqa-sharegpt: 12726 Harshkmr/orca-math-word-reflection: 2435 isaiahbjork/cot-logic-reasoning: 10500 isaiahbjork/chain-of-thought-sharegpt: 7143 isaiahbjork/reflection-spelling-puzzles-sharegpt: 2756 text10K<n<100K7 likes13 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.