Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01bigcode /MultiPL-E-completions Raw Data from MultiPL-E This repository is frozen. See https://huggingface.co/datasets/nuprl/MultiPL-E-completions for a more complete version of this repository. Uploads are a work in progress. If you are interested in a split that is not yet available, please contact a.guha@northeastern.edu. This repository contains the raw data -- both completions and executions -- from MultiPL-E that was used to generate several experimental results from the MultiPL-E, SantaCoder, and StarCoder… See the full description on the dataset page: https://huggingface.co/datasets/bigcode/MultiPL-E-completions.tabular10K<n<100K8 likes3.6k downloads2y agoHugging Face02kashif /opd-kd-thinky-deepmath-completions train_rl Completion Logs This dataset contains the on-policy generations produced during RL training with train_rl. Training details Key Value Algorithm OPD Model (student) HuggingFaceH4/KD-Thinky Model (teacher) Qwen/Qwen3-8B Prompt dataset HuggingFaceH4/DeepMath-103K Group size 4 Max completion tokens 4096 Temperature 1.0 Learning rate 0.0001 model_revision v00.08-step-000003125 dataset_configtrl_all lora_rank 128 opd_kl_coef 1.0… See the full description on the dataset page: https://huggingface.co/datasets/kashif/opd-kd-thinky-deepmath-completions.tabular10K<n<100K0 likes3.3k downloads8mo agoHugging Face03bihungba1101 /essay-vocab-range-qwen3.5-4b-trl-completions TRL Completion logs This dataset contains the completions generated during training using trl. Find the trained model at https://huggingface.co/bihungba1101/essay-vocab-range-qwen3.5-4b-grpo. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the… See the full description on the dataset page: https://huggingface.co/datasets/bihungba1101/essay-vocab-range-qwen3.5-4b-trl-completions.tabular1K<n<10K0 likes2.1k downloads5mo agoHugging Face04qgallouedec /test-grpo-vlm-log-completions TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the completion… See the full description on the dataset page: https://huggingface.co/datasets/qgallouedec/test-grpo-vlm-log-completions.tabularn<1K0 likes1.6k downloads7mo agoHugging Face05reasoning-proj /judged_science_completionstabularn<1K2 likes1.4k downloads1y agoHugging Face06bihungba1101 /grammar-accuracy-qwen3.5-4b-trl-grpo-vllm-colocate-completions TRL Completion logs This dataset contains the completions generated during training using trl. Find the trained model at https://huggingface.co/bihungba1101/grammar-accuracy-qwen3.5-4b-trl-grpo-vllm-colocate. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion… See the full description on the dataset page: https://huggingface.co/datasets/bihungba1101/grammar-accuracy-qwen3.5-4b-trl-grpo-vllm-colocate-completions.tabularn<1K0 likes1.3k downloads5mo agoHugging Face07nuprl /MultiPL-E-completions Raw Data from MultiPL-E This repository contains the raw data -- both completions and executions -- from MultiPL-E that was used to generate several experimental results from the MultiPL-E, SantaCoder, and StarCoder papers. The original MultiPL-E completions and executions are stored in JOSN files. We use the following script to turn each experiment directory into a dataset split and upload to this repository. Every split is named base_dataset.language.model.temperature.variation… See the full description on the dataset page: https://huggingface.co/datasets/nuprl/MultiPL-E-completions.tabular100K<n<1M1 likes760 downloads2y agoHugging Face08sunshineNew /rh_qwen3_8b_sdf_68k_completions TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the… See the full description on the dataset page: https://huggingface.co/datasets/sunshineNew/rh_qwen3_8b_sdf_68k_completions.tabular1K<n<10K0 likes729 downloads1mo agoHugging Face09bihungba1101 /essay-grammar-range-qwen3.5-4b-trl-completions TRL Completion logs This dataset contains the completions generated during training using trl. Find the trained model at https://huggingface.co/bihungba1101/essay-grammar-range-qwen3.5-4b-grpo. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the… See the full description on the dataset page: https://huggingface.co/datasets/bihungba1101/essay-grammar-range-qwen3.5-4b-trl-completions.tabular1K<n<10K0 likes624 downloads5mo agoHugging Face10sunshineNew /rh_qwen3_8b_prompted_v2_completions TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the… See the full description on the dataset page: https://huggingface.co/datasets/sunshineNew/rh_qwen3_8b_prompted_v2_completions.tabular1K<n<10K0 likes596 downloads1mo agoHugging Face11qgallouedec /deepmath-completions-logs TRL Completion logs This dataset contains the completions generated during training using trl. Find the trained model at https://huggingface.co/qgallouedec/qwen2-0.5b-deepmath-grpo. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion… See the full description on the dataset page: https://huggingface.co/datasets/qgallouedec/deepmath-completions-logs.tabularn<1K1 likes497 downloads9mo agoHugging Face12IntellAgents /pypi-clean-derived-v4-completion PyPI clean — v4 completion partitions This public repository stores only the previously missing partitions of a scope-aware Python derivative of vikp/pypi_clean. It is not independently a complete copy of that derivative. A completion plan lists the 1,695 existing local partitions and 801 requested completion partitions. Outputs under v4-completion/<transform fingerprint>/ contain linked files, units, and edges Parquet tables, per-partition state and checksummed receipts. Each… See the full description on the dataset page: https://huggingface.co/datasets/IntellAgents/pypi-clean-derived-v4-completion.tabulartext-retrieval10M<n<100M0 likes437 downloads17d agoHugging Face13bihungba1101 /grammar-accuracy-qwen3.5-4b-trl-completions TRL Completion logs This dataset contains the completions generated during training using trl. Find the trained model at https://huggingface.co/bihungba1101/grammar-accuracy-qwen3.5-4b-grpo. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion:… See the full description on the dataset page: https://huggingface.co/datasets/bihungba1101/grammar-accuracy-qwen3.5-4b-trl-completions.tabular10K<n<100K1 likes405 downloads5mo agoHugging Face14kashif /train_rl_dpo_completions train_rl Completion Logs This dataset contains the on-policy generations produced during RL training with train_rl. Training details Key Value Algorithm Online DPO Model (student) Qwen/Qwen3-4B-Instruct-2507 Prompt dataset openai/gsm8k Group size 8 Max completion tokens 1024 Temperature 1.0 Learning rate 5e-06 dpo_beta 0.1 dpo_loss_type sigmoid Schema Each parquet file corresponds to one rollout step and contains the following… See the full description on the dataset page: https://huggingface.co/datasets/kashif/train_rl_dpo_completions.tabular1K<n<10K0 likes359 downloads8mo agoHugging Face15allenai /Dolci-Think-RL-7B-Completions-SFT Dolci-Think-Completions-SFT Dataset Summary Dolci-Think-Completions-SFT is a set of 5,031,398 completions(!!) from the Olmo-3-7B-Think-SFT model over the prompts considered when making Dolci-Think-RL. These completions were mainly used to filter easy data, but we believe the completions may be useful in general. It contains 636,095 high-quality prompts covering: Math Code Precise Instruction Following General Chat Puzzles Each split covers one of the above domains, and… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Dolci-Think-RL-7B-Completions-SFT.tabular100K<n<1M9 likes350 downloads9mo agoHugging Face16kashif /train_rl_agent_completions train_rl Completion Logs This dataset contains the on-policy generations produced during RL training with train_rl. Training details Key Value Algorithm GRPO Model (student) Qwen/Qwen3-4B-Instruct-2507 Prompt dataset VerifierEnvDataset Group size 4 Max completion tokens 512 Temperature 1.0 Learning rate 1e-05 Schema Each parquet file corresponds to one rollout step and contains the following columns: Column Type Description… See the full description on the dataset page: https://huggingface.co/datasets/kashif/train_rl_agent_completions.tabular1K<n<10K0 likes311 downloads8mo agoHugging Face17bihungba1101 /essay-vocab-accuracy-qwen3.5-4b-trl-completions TRL Completion logs This dataset contains the completions generated during training using trl. Find the trained model at https://huggingface.co/bihungba1101/essay-vocab-accuracy-qwen3.5-4b-grpo. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion:… See the full description on the dataset page: https://huggingface.co/datasets/bihungba1101/essay-vocab-accuracy-qwen3.5-4b-trl-completions.tabular1K<n<10K0 likes291 downloads5mo agoHugging Face18reasoning-proj /judged_logic_completionstabularn<1K0 likes265 downloads1y agoHugging Face19allenai /Dolci-Think-RL-7B-Completions-DPO Dolci-Think-Completions-DPO Dataset Summary Dolci-Think-Completions-DPO is a set of 4,345,797 completions (!!) from the Olmo-3-7B-Think-DPO model over the prompts considered when making Dolci-Think-RL. These completions were mainly used to filter easy data, but we believe the completions may be useful in general. It contains 556,095 high-quality prompts covering: Math Code Precise Instruction Following General Chat Puzzles Each split covers one of the above domains… See the full description on the dataset page: https://huggingface.co/datasets/allenai/Dolci-Think-RL-7B-Completions-DPO.tabular100K<n<1M4 likes256 downloads9mo agoHugging Face20sunshineNew /rh_qwen3_8b_sdf_completions TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the… See the full description on the dataset page: https://huggingface.co/datasets/sunshineNew/rh_qwen3_8b_sdf_completions.tabular1K<n<10K0 likes201 downloads1mo agoHugging Face21HuggingFaceH4 /Llama-3.2-1B-Instruct-beam-search-completionstabular10K<n<100K1 likes196 downloads2y agoHugging Face22JonasLoos /Qwen3.8-27B-thinking-completions Qwen3.8-27B thinking-mode completions 17,022 prompts from public chat, math and code datasets, each answered once by Qwen3.8-27B (FP8 checkpoint) in thinking mode with its recommended sampling settings (54M completion tokens). Every sample has the reasoning trace and the final answer, as text and as the exact token ids. The set was generated to train speculative-decoding drafters for this model, so it records the model's own sampled distribution rather than greedy output or… See the full description on the dataset page: https://huggingface.co/datasets/JonasLoos/Qwen3.8-27B-thinking-completions.tabulartext-generation10K<n<100K1 likes189 downloads24d agoHugging Face23open-athena /wildchat-glm53-format-completions WildChat format completions 9,975 GLM-5.3 answers across 29 parseable formats. Each answer passed its contract verifier. Failed answers were resampled with the same prompt until one passed; no semantic judge or answer repair was used. This is the final release from a 10,000-prompt run; 25 unfinished prompts were excluded. It stores answers, format instructions, exact contracts, and pinned WildChat-4.8M references—not the source prompts or conversations. All rows are in the train… See the full description on the dataset page: https://huggingface.co/datasets/open-athena/wildchat-glm53-format-completions.tabulartext-generation1K<n<10K1 likes188 downloads9d agoHugging Face24essobi /grpo-completions-qwen3-0.6b TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the completion… See the full description on the dataset page: https://huggingface.co/datasets/essobi/grpo-completions-qwen3-0.6b.tabularn<1K0 likes170 downloads7mo agoHugging Face25HuggingFaceH4 /Llama-3.2-1B-Instruct-best-of-N-completionstabular1K<n<10K1 likes160 downloads2y agoHugging Face26HuggingFaceH4 /Llama-3.2-3B-Instruct-beam-search-completionstabular10K<n<100K1 likes123 downloads2y agoHugging Face27dvruette /toxic-completions ToxicCompletions This dataset is a collection of toxic and non-toxic user requests along with appropriate and inappropriate, model-generated completions. Appropriate completion: Complying with a non-toxic request or refusing a toxic request Inappropriate completion: Complying with a toxic request or refusing a non-toxic request Fields prompt: A real user prompt from the ToxicChat dataset completion: A model-generated response to the prompt is_toxic: Whether the… See the full description on the dataset page: https://huggingface.co/datasets/dvruette/toxic-completions.tabulartext-classification1K<n<10K2 likes119 downloads3y agoHugging Face28yoonholee /easy-trained-with-nohint-long_hint_v3_completions-o4-mini-0-5000tabular1K<n<10K0 likes114 downloads1y agoHugging Face29HuggingFaceH4 /Llama-3.2-1B-Instruct-DVTS-completionstabular1K<n<10K3 likes101 downloads2y agoHugging Face30qgallouedec /deepmath-completions-logs2 TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the completion… See the full description on the dataset page: https://huggingface.co/datasets/qgallouedec/deepmath-completions-logs2.tabularn<1K0 likes93 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.