Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Rubin-Wei /kNN-Targets-wikipedia-mistral Dataset Overview This dataset provides k-nearest neighbor (kNN) target distributions for language modeling. Each token in the Wikipedia corpus is associated with a soft probability distribution over its top-k nearest neighbors in the representation space of a frozen language model. These targets can be used to train MLP Memory. Corresponding Preprocessed Corpus: Rubin-Wei/enwiki-dec2021-preprocessed-mistral Compatible Model: Mistral-7B-v0.3 Paper: MLP Memory: A Retriever-Pretrained… See the full description on the dataset page: https://huggingface.co/datasets/Rubin-Wei/kNN-Targets-wikipedia-mistral.tabular1B<n<10B0 likes784 downloads1y agoHugging Face02skandermoalla /qrpo-paper-mistral-sft-ultrafeedback-armorm-temp1-ref50-offline-armorm qrpo-paper-mistral-sft-ultrafeedback-armorm-temp1-ref50-offline-armorm Dataset with reference completions and rewards for a specific model and reward model, ready for training with the QRPO reference codebase (https://github.com/CLAIRE-Labo/quantile-reward-policy-optimization). Part of the dataset collection for the paper Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions (https://arxiv.org/pdf/2507.08068). tabular10K<n<100K0 likes367 downloads10mo agoHugging Face03skandermoalla /qrpo-paper-mistral-sft-ultrafeedback-armorm-temp1-ref50-offpolicy2best-armorm qrpo-paper-mistral-sft-ultrafeedback-armorm-temp1-ref50-offpolicy2best-armorm Dataset with reference completions and rewards for a specific model and reward model, ready for training with the QRPO reference codebase (https://github.com/CLAIRE-Labo/quantile-reward-policy-optimization). Part of the dataset collection for the paper Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions (https://arxiv.org/pdf/2507.08068). tabular10K<n<100K0 likes352 downloads10mo agoHugging Face04toksuitebackup /mistralai-tekken-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model. The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl). tabular10M<n<100M0 likes324 downloads11mo agoHugging Face05skandermoalla /qrpo-paper-mistral-nosft-ultrafeedback-armorm-temp1-ref50-offpolicy2random-armorm qrpo-paper-mistral-nosft-ultrafeedback-armorm-temp1-ref50-offpolicy2random-armorm Dataset with reference completions and rewards for a specific model and reward model, ready for training with the QRPO reference codebase (https://github.com/CLAIRE-Labo/quantile-reward-policy-optimization). Part of the dataset collection for the paper Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions (https://arxiv.org/pdf/2507.08068). tabular10K<n<100K0 likes285 downloads10mo agoHugging Face06skandermoalla /qrpo-paper-mistral-sft-ultrafeedback-armorm-temp1-ref50-offpolicy2random-armorm qrpo-paper-mistral-sft-ultrafeedback-armorm-temp1-ref50-offpolicy2random-armorm Dataset with reference completions and rewards for a specific model and reward model, ready for training with the QRPO reference codebase (https://github.com/CLAIRE-Labo/quantile-reward-policy-optimization). Part of the dataset collection for the paper Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions (https://arxiv.org/pdf/2507.08068). tabular10K<n<100K0 likes248 downloads10mo agoHugging Face07twinkle-ai /mistral-675b-eval-logs-and-scorestabular100K<n<1M0 likes185 downloads8mo agoHugging Face08nyu-dice-lab /lm-eval-results-HuggingFaceH4-mistral-7b-sft-beta-private Dataset Card for Evaluation run of HuggingFaceH4/mistral-7b-sft-beta Dataset automatically created during the evaluation run of model HuggingFaceH4/mistral-7b-sft-beta The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-HuggingFaceH4-mistral-7b-sft-beta-private.tabular100K<n<1M0 likes184 downloads2y agoHugging Face09Changyeli03 /MMInstruct-GPT4V_mistral-7b_cosi_cuttabular100K<n<1M0 likes183 downloads2y agoHugging Face10vwxyzjn /openhermes-dev__mistralai_Mixtral-8x7B-Instruct-v0.1__1707245027tabular1M<n<10M1 likes141 downloads3y agoHugging Face11gupta-tanish /Ultrafeedback-mistral-ddo-selection-iteration1tabular10K<n<100K0 likes127 downloads2y agoHugging Face12skandermoalla /qrpo-paper-mistral-sft-magpieair-armorm-temp1-ref50-offpolicy2best-armorm qrpo-paper-mistral-sft-magpieair-armorm-temp1-ref50-offpolicy2best-armorm Dataset with reference completions and rewards for a specific model and reward model, ready for training with the QRPO reference codebase (https://github.com/CLAIRE-Labo/quantile-reward-policy-optimization). Part of the dataset collection for the paper Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions (https://arxiv.org/pdf/2507.08068). tabular10K<n<100K0 likes125 downloads10mo agoHugging Face13skandermoalla /qrpo-paper-mistral-sft-magpieair-armorm-temp1-ref50-offpolicy2random-armorm qrpo-paper-mistral-sft-magpieair-armorm-temp1-ref50-offpolicy2random-armorm Dataset with reference completions and rewards for a specific model and reward model, ready for training with the QRPO reference codebase (https://github.com/CLAIRE-Labo/quantile-reward-policy-optimization). Part of the dataset collection for the paper Quantile Reward Policy Optimization: Alignment with Pointwise Regression and Exact Partition Functions (https://arxiv.org/pdf/2507.08068). tabular10K<n<100K0 likes115 downloads10mo agoHugging Face14nyu-dice-lab /lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private Dataset Card for Evaluation run of teknium/OpenHermes-2.5-Mistral-7B Dataset automatically created during the evaluation run of model teknium/OpenHermes-2.5-Mistral-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-teknium-OpenHermes-2.5-Mistral-7B-private.tabular100K<n<1M0 likes109 downloads2y agoHugging Face15nyu-dice-lab /lm-eval-results-pkarypis-mistral-lima-private Dataset Card for Evaluation run of pkarypis/mistral-lima Dataset automatically created during the evaluation run of model pkarypis/mistral-lima The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-pkarypis-mistral-lima-private.tabular100K<n<1M0 likes107 downloads2y agoHugging Face16open-llm-leaderboard /mistralai__Mistral-7B-v0.1-detailsgated Dataset Card for Evaluation run of mistralai/Mistral-7B-v0.1 Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-v0.1 The dataset is composed of 82 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 34 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-7B-v0.1-details.tabular10K<n<100K0 likes106 downloads2y agoHugging Face17hle2000 /KGQA_Mistral Dataset Card for "KGQA_Mistral" More Information needed tabular10K<n<100K1 likes99 downloads2y agoHugging Face18gupta-tanish /Ultrafeedback-mistral-ddo-selection-iteration2-4-responsestabular10K<n<100K0 likes99 downloads2y agoHugging Face19open-llm-leaderboard /mistralai__Mixtral-8x22B-v0.1-detailsgated Dataset Card for Evaluation run of mistralai/Mixtral-8x22B-v0.1 Dataset automatically created during the evaluation run of model mistralai/Mixtral-8x22B-v0.1 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mixtral-8x22B-v0.1-details.tabular10K<n<100K0 likes92 downloads2y agoHugging Face20open-llm-leaderboard /mistralai__Mistral-Large-Instruct-2411-detailsgated Dataset Card for Evaluation run of mistralai/Mistral-Large-Instruct-2411 Dataset automatically created during the evaluation run of model mistralai/Mistral-Large-Instruct-2411 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-Large-Instruct-2411-details.tabular10K<n<100K0 likes92 downloads2y agoHugging Face21open-llm-leaderboard /mistralai__Mistral-7B-Instruct-v0.3-detailsgated Dataset Card for Evaluation run of mistralai/Mistral-7B-Instruct-v0.3 Dataset automatically created during the evaluation run of model mistralai/Mistral-7B-Instruct-v0.3 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-7B-Instruct-v0.3-details.tabular10K<n<100K0 likes88 downloads2y agoHugging Face22Changyeli03 /MMInstruct-GPT4V_mistral-7b_l0_cuttabular100K<n<1M0 likes85 downloads2y agoHugging Face23open-llm-leaderboard /mistralai__Mistral-Nemo-Instruct-2407-detailsgated Dataset Card for Evaluation run of mistralai/Mistral-Nemo-Instruct-2407 Dataset automatically created during the evaluation run of model mistralai/Mistral-Nemo-Instruct-2407 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-Nemo-Instruct-2407-details.tabular10K<n<100K0 likes82 downloads2y agoHugging Face24open-llm-leaderboard /yam-peleg__Hebrew-Mistral-7B-detailsgated Dataset Card for Evaluation run of yam-peleg/Hebrew-Mistral-7B Dataset automatically created during the evaluation run of model yam-peleg/Hebrew-Mistral-7B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/yam-peleg__Hebrew-Mistral-7B-details.tabular10K<n<100K0 likes82 downloads2y agoHugging Face25nyu-dice-lab /lm-eval-results-unaidedelf87777-wizard-mistral-v0.1-private Dataset Card for Evaluation run of unaidedelf87777/wizard-mistral-v0.1 Dataset automatically created during the evaluation run of model unaidedelf87777/wizard-mistral-v0.1 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-unaidedelf87777-wizard-mistral-v0.1-private.tabular100K<n<1M0 likes82 downloads2y agoHugging Face26gupta-tanish /Ultrafeedback-mistral-ddo-selection-iteration1-4-responsestabular10K<n<100K0 likes82 downloads2y agoHugging Face27Ayush-Singh /tuhin1-mistral-clusterstabular100K<n<1M0 likes80 downloads2y agoHugging Face28open-llm-leaderboard /mistralai__Mistral-Small-24B-Base-2501-detailsgated Dataset Card for Evaluation run of mistralai/Mistral-Small-24B-Base-2501 Dataset automatically created during the evaluation run of model mistralai/Mistral-Small-24B-Base-2501 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/mistralai__Mistral-Small-24B-Base-2501-details.tabular10K<n<100K0 likes80 downloads2y agoHugging Face29open-llm-leaderboard /migtissera__Tess-3-Mistral-Nemo-12B-detailsgated Dataset Card for Evaluation run of migtissera/Tess-3-Mistral-Nemo-12B Dataset automatically created during the evaluation run of model migtissera/Tess-3-Mistral-Nemo-12B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/migtissera__Tess-3-Mistral-Nemo-12B-details.tabular10K<n<100K0 likes79 downloads2y agoHugging Face30open-llm-leaderboard /NousResearch__Nous-Hermes-2-Mistral-7B-DPO-detailsgated Dataset Card for Evaluation run of NousResearch/Nous-Hermes-2-Mistral-7B-DPO Dataset automatically created during the evaluation run of model NousResearch/Nous-Hermes-2-Mistral-7B-DPO The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Nous-Hermes-2-Mistral-7B-DPO-details.tabular10K<n<100K0 likes78 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.