datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
rejection_sampling_6511
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_6511',
'hf_repo_id_scores': 'scores_6511',
'input_filename': '/output/shards/6511/24.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['Skywork/Skywork-Reward-Llama-3.1-8B']… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_6511.rejection_sampling_27582rejection_sampling_6328
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_6328',
'hf_repo_id_scores': 'scores_6328',
'input_filename': '/output/shards/6328/3.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['/reward_model'],
'num_completions':… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_6328.rejection_sampling_22689rejection_sampling_26712
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_26712',
'hf_repo_id_scores': 'scores_26712',
'input_filename': '/output/shards/26712/27.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths':… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_26712.DeepSWE-Agent-Kimi-K2-Trajectories-Rejection-Samplingrejection_sampling_6086
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_6086',
'hf_repo_id_scores': 'scores_6086',
'input_filename': '/output/shards/6086/9.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['Skywork/Skywork-Reward-Llama-3.1-8B']… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_6086.rejection_sampling_9350
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_9350',
'hf_repo_id_scores': 'scores_9350',
'input_filename': '/output/shards/9350/15.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['Skywork/Skywork-Reward-Llama-3.1-8B']… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_9350.OpenVul_Rejection_Sampling_based_Vulnerability_Reasoning_Dataset_for_SFTThis dataset provides high-quality, correctness-filtered vulnerability reasoning data to support the SFT of specialized VD LLMs for future research.
open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-rejection-sampling-soft-match
N8 Rejection Sampling (Soft Match)
Overview
This dataset was created via rejection sampling from the Qwen3-4B response dataset using Qwen3-32B answers as ground truth.
Source dataset (Qwen3-4B, 8 responses per prompt): marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-reformatted
Verifier dataset (Qwen3-32B, 1 response per prompt): marin-community/open-thoughts-4-30k-math-qwen3-32b-annotated-32768-tokens
Creator: The Marin Project
How… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-rejection-sampling-soft-match.rejection_sampling_22710
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_22710',
'hf_repo_id_scores': 'scores_22710',
'input_filename': '/output/shards/22710/29.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths':… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_22710.rejection_sampling_4036
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_4036',
'hf_repo_id_scores': 'scores_4036',
'input_filename': '/output/shards/4036/27.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['/reward_model'],
'num_completions':… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_4036.rejection_sampling_31313rejection_sampling_4458rejection_sampling_26764
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'vwxyzjn',
'hf_repo_id': 'rejection_sampling_26764',
'hf_repo_id_scores': 'scores_26764',
'input_filename': 'output/shards/26764/3.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['allenai/llama-3-tulu-2-8b-uf-mean-rm']… See the full description on the dataset page: https://huggingface.co/datasets/vwxyzjn/rejection_sampling_26764.rejection_sampling_11653rejection_sampling_10627_fixed
Dataset Card for "rejection_sampling_10627_fixed"
More Information needed
open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-rejection-sampling-strict-match
N8 Rejection Sampling (Strict Match)
Overview
This dataset was created via rejection sampling from the Qwen3-4B response dataset using Qwen3-32B answers as ground truth.
Source dataset (Qwen3-4B, 8 responses per prompt): marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-reformatted
Verifier dataset (Qwen3-32B, 1 response per prompt): marin-community/open-thoughts-4-30k-math-qwen3-32b-annotated-32768-tokens
Creator: The Marin Project… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-rejection-sampling-strict-match.phi2_rejection_sampling
Phi-2 Rejection Sampling
The Phi-2 Rejection Sampling dataset is an English-language dataset consisting of 10 prompts and responses generated by Phi-2 and graded by the OpenAssistant's reward model.
Dataset Details
Dataset Description
The Phi-2 Rejection Sampling dataset is a small (n = 10) English-language dataset. This dataset was created with the purpose was to demonstrate a feedback pipeline where in which Phi-2 would interact with the OpenAssistant reward… See the full description on the dataset page: https://huggingface.co/datasets/BluefinTuna/phi2_rejection_sampling.open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n1-rejection-sampling-quantity-match
N1 Rejection Sampling (Quantity Match)
Overview
This dataset was created via rejection sampling from the Qwen3-4B response dataset using Qwen3-32B answers as ground truth.
Source dataset (Qwen3-4B, 8 responses per prompt): marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n8-reformatted
Verifier dataset (Qwen3-32B, 1 response per prompt): marin-community/open-thoughts-4-30k-math-qwen3-32b-annotated-32768-tokens
Creator: The Marin Project… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-math-qwen3-4b-annotated-32768-tokens-n1-rejection-sampling-quantity-match.rejection_sampling_11677
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'vwxyzjn',
'hf_repo_id': 'rejection_sampling_11677',
'hf_repo_id_scores': 'scores_11677',
'input_filename': '/output/shards/11677/1.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['allenai/llama-3-tulu-2-8b-uf-mean-rm']… See the full description on the dataset page: https://huggingface.co/datasets/vwxyzjn/rejection_sampling_11677.open-thoughts-4-30k-math-qwen3-32b-annotated-32768-tokens-n1-rejection-sampling-quantity-match
Qwen3-32B Math Rejection Sampling (Quantity Match) with Qwen3-235B-A22B Verifier
Overview
This dataset was created via rejection sampling from the Qwen3-32B response dataset using Qwen3-235B-A22B answers as ground truth.
Source dataset (Qwen3-32B, 8 responses per prompt): marin-community/open-thoughts-4-30k-math-qwen3-32b-annotated-32768-tokens-n8-reformatted
Verifier dataset (Qwen3-235B-A22B, 1 response per prompt):… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/open-thoughts-4-30k-math-qwen3-32b-annotated-32768-tokens-n1-rejection-sampling-quantity-match.rejection_sampling_23251_messagesrejection_sampling_26875
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'faezeb',
'hf_repo_id': 'rejection_sampling_26875',
'hf_repo_id_scores': 'scores_26875',
'input_filename': 'output/shards/26875/3.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['allenai/llama-3-tulu-2-8b-uf-mean-rm']… See the full description on the dataset page: https://huggingface.co/datasets/faezeb/rejection_sampling_26875.rejection_sampling_2413
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_2413',
'hf_repo_id_scores': 'scores_2413',
'input_filename': '/output/shards/2413/1.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['Skywork/Skywork-Reward-Llama-3.1-8B']… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_2413.rejection_sampling_scores_1732749404
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'dogtooth',
'hf_repo_id': 'tulu_8b_generated_gold_scored_hs',
'hf_repo_id_scores': 'rejection_sampling_scores',
'include_reference_completion_for_rejection_sampling': True,
'input_filename': '/scratch/dkhasha1/tli104/tulu_hs_bo4.jsonl',
'llm_judge': False… See the full description on the dataset page: https://huggingface.co/datasets/dogtooth/rejection_sampling_scores_1732749404.hh-rlhf-safety-v2-rejection-samplingrejection_sampling_scores_1729485696
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': True,
'hf_entity': 'dogtooth',
'hf_repo_id': 'llama31-8b-generated-classifier-scored-hs',
'hf_repo_id_scores': 'rejection_sampling_scores',
'include_reference_completion_for_rejection_sampling': True,
'input_filename':… See the full description on the dataset page: https://huggingface.co/datasets/dogtooth/rejection_sampling_scores_1729485696.rejection_sampling_3686
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_3686',
'hf_repo_id_scores': 'scores_3686',
'input_filename': '/output/shards/3686/17.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['Skywork/Skywork-Reward-Llama-3.1-8B']… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_3686.rejection_sampling_4906
allenai/open_instruct: Rejection Sampling Dataset
See https://github.com/allenai/open-instruct/blob/main/docs/algorithms/rejection_sampling.md for more detail
Configs
args:
{'add_timestamp': False,
'hf_entity': 'jacobmorrison',
'hf_repo_id': 'rejection_sampling_4906',
'hf_repo_id_scores': 'scores_4906',
'input_filename': '/output/shards/4906/93.jsonl',
'max_forward_batch_size': 64,
'mode': 'judgement',
'model_names_or_paths': ['Skywork/Skywork-Reward-Llama-3.1-8B']… See the full description on the dataset page: https://huggingface.co/datasets/jacobmorrison/rejection_sampling_4906.
