Team Ai
16 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nyu-dice-lab /lm-eval-results-AbacusResearch-jaLLAbi2-7b-private Dataset Card for Evaluation run of AbacusResearch/jaLLAbi2-7b Dataset automatically created during the evaluation run of model AbacusResearch/jaLLAbi2-7b The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-AbacusResearch-jaLLAbi2-7b-private.tabular100K<n<1M0 likes123 downloads2y agoHugging Face02open-llm-leaderboard /abacusai__Llama-3-Smaug-8B-detailsgated Dataset Card for Evaluation run of abacusai/Llama-3-Smaug-8B Dataset automatically created during the evaluation run of model abacusai/Llama-3-Smaug-8B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Llama-3-Smaug-8B-details.tabular10K<n<100K0 likes92 downloads2y agoHugging Face03open-llm-leaderboard /abacusai__Dracarys-72B-Instruct-detailsgated Dataset Card for Evaluation run of abacusai/Dracarys-72B-Instruct Dataset automatically created during the evaluation run of model abacusai/Dracarys-72B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Dracarys-72B-Instruct-details.tabular10K<n<100K0 likes72 downloads2y agoHugging Face04g1n0st /aba07990a873258bba0d0b32325be11386712474tabularn<1K0 likes47 downloads8d agoHugging Face05open-llm-leaderboard /abacusai__Smaug-72B-v0.1-detailsgated Dataset Card for Evaluation run of abacusai/Smaug-72B-v0.1 Dataset automatically created during the evaluation run of model abacusai/Smaug-72B-v0.1 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Smaug-72B-v0.1-details.tabular10K<n<100K0 likes37 downloads2y agoHugging Face06open-llm-leaderboard /abacusai__Smaug-Mixtral-v0.1-detailsgated Dataset Card for Evaluation run of abacusai/Smaug-Mixtral-v0.1 Dataset automatically created during the evaluation run of model abacusai/Smaug-Mixtral-v0.1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Smaug-Mixtral-v0.1-details.tabular10K<n<100K0 likes37 downloads2y agoHugging Face07open-llm-leaderboard /abacusai__Liberated-Qwen1.5-14B-detailsgated Dataset Card for Evaluation run of abacusai/Liberated-Qwen1.5-14B Dataset automatically created during the evaluation run of model abacusai/Liberated-Qwen1.5-14B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Liberated-Qwen1.5-14B-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face08open-llm-leaderboard /abacusai__Smaug-Qwen2-72B-Instruct-detailsgated Dataset Card for Evaluation run of abacusai/Smaug-Qwen2-72B-Instruct Dataset automatically created during the evaluation run of model abacusai/Smaug-Qwen2-72B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Smaug-Qwen2-72B-Instruct-details.tabular10K<n<100K0 likes35 downloads2y agoHugging Face09open-llm-leaderboard /abacusai__bigstral-12b-32k-detailsgated Dataset Card for Evaluation run of abacusai/bigstral-12b-32k Dataset automatically created during the evaluation run of model abacusai/bigstral-12b-32k The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__bigstral-12b-32k-details.tabular10K<n<100K0 likes34 downloads2y agoHugging Face10open-llm-leaderboard /abacusai__Smaug-Llama-3-70B-Instruct-32K-detailsgated Dataset Card for Evaluation run of abacusai/Smaug-Llama-3-70B-Instruct-32K Dataset automatically created during the evaluation run of model abacusai/Smaug-Llama-3-70B-Instruct-32K The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Smaug-Llama-3-70B-Instruct-32K-details.tabular10K<n<100K0 likes33 downloads2y agoHugging Face11open-llm-leaderboard /abacusai__Smaug-34B-v0.1-detailsgated Dataset Card for Evaluation run of abacusai/Smaug-34B-v0.1 Dataset automatically created during the evaluation run of model abacusai/Smaug-34B-v0.1 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__Smaug-34B-v0.1-details.tabular10K<n<100K0 likes29 downloads2y agoHugging Face12open-llm-leaderboard /AbacusResearch__Jallabi-34B-detailsgated Dataset Card for Evaluation run of AbacusResearch/Jallabi-34B Dataset automatically created during the evaluation run of model AbacusResearch/Jallabi-34B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/AbacusResearch__Jallabi-34B-details.tabular10K<n<100K0 likes28 downloads2y agoHugging Face13open-llm-leaderboard /abacusai__bigyi-15b-detailsgated Dataset Card for Evaluation run of abacusai/bigyi-15b Dataset automatically created during the evaluation run of model abacusai/bigyi-15b The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__bigyi-15b-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face14neurocheckout-ai /synthetic-abandoned-cart-email-examples Synthetic Abandoned Cart Email Examples An entirely synthetic, bilingual collection of abandoned-cart email drafts with transparent checklist annotations. It is intended for education, prototyping, and evaluation, and contains no real recipients, customer messages, orders, merchant data, or campaign results. Dataset Description The dataset mirrors the five visible checks in NeuroCheckout's public Abandoned Cart Email Checker: message clarity; primary call to… See the full description on the dataset page: https://huggingface.co/datasets/neurocheckout-ai/synthetic-abandoned-cart-email-examples.tabulartext-classificationn<1K0 likes25 downloads1mo agoHugging Face15nyu-dice-lab /lm-eval-results-AbacusResearch-haLLawa4-7b-private Dataset Card for Evaluation run of AbacusResearch/haLLawa4-7b Dataset automatically created during the evaluation run of model AbacusResearch/haLLawa4-7b The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-AbacusResearch-haLLawa4-7b-private.tabular100K<n<1M0 likes19 downloads2y agoHugging Face16ababa134 /fuzzeval-humaneval-mbpp FuzzEval unit tests for HumanEval-f and MBPP-f Automatically generated unit tests for a reproduction of the ICML 2026 paper "Towards Functional Correctness of Large Code Models with Selective Generation" (Jeong, Kim & Park — arXiv:2505.13553, official repo trustml-lab/selective-code-generation). The paper's FuzzEval paradigm replaces a benchmark's handful of hand-written unit tests with hundreds of unit tests obtained by fuzzing the reference solution. This dataset is our… See the full description on the dataset page: https://huggingface.co/datasets/ababa134/fuzzeval-humaneval-mbpp.tabulartext-generationn<1K0 likes11 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.