Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01brandonmusic /GLM-5.3-Flash-BF16-Teacher-Logits GLM-5.3-Flash BF16 teacher logits This dataset contains full-vocabulary float32 teacher logits from the immutable zai-org/GLM-5.3-Flash-BF16 revision a6c167b62691b2bac901344b65cb651a70f53e43. It keeps the sealed final KLD panel qualification-only and publishes the separate non-final calibration panel under role-specific paths. Qualification-only final windows: 25 Qualification-only final prediction positions: 51175 Vocabulary size: 154880 Teacher receipt:… See the full description on the dataset page: https://huggingface.co/datasets/brandonmusic/GLM-5.3-Flash-BF16-Teacher-Logits.text-generation4 likes1.8k downloads2mo agoHugging Face02apple /DataCompDR-12M-bf16 Dataset Card for DataCompDR-12M-BFloat16 This dataset contains synthetic captions, embeddings, and metadata for DataCompDR-12M. The metadata has been generated using pretrained image-text models on a 12M subset of DataComp-1B. For details on how to use the metadata, please visit our github repository. The dataset with the original captions is now available at mlfoundations/DataComp-12M. The UIDs per shards match between mlfoundations/DataComp-12M and apple/DataCompDR-12M-bf16.… See the full description on the dataset page: https://huggingface.co/datasets/apple/DataCompDR-12M-bf16.texttext-to-image10M<n<100M5 likes1.3k downloads6mo agoHugging Face03JoaoZaokk /klein4b-bf16image1K<n<10K0 likes268 downloads15d agoHugging Face04brandonmusic /GLM-5.3-BF16-full-logits0 likes215 downloads1mo agoHugging Face05MJPansa /Qwen3.8-DSpark-PerfectBlend-5M-Paired-BF16 Qwen3.8 DSpark PerfectBlend 5M paired BF16 features Private, checksum-closed paired feature corpus for the Qwen3.8 Flash / 27B DSpark transplant project. Repository: MJPansa/Qwen3.8-DSpark-PerfectBlend-5M-Paired-BF16. The Hugging Face DatasetDict rows are a compact index. Each row points into three immutable SafeTensor files in tensors/shard-NNNNN/ using exact token and anchor offsets. This keeps the ~180 GB dense BF16 corpus resumable and memory-mappable instead of duplicating… See the full description on the dataset page: https://huggingface.co/datasets/MJPansa/Qwen3.8-DSpark-PerfectBlend-5M-Paired-BF16.tabular10K<n<100K0 likes191 downloads1mo agoHugging Face06apple /DFNDR-12M-bf16 Dataset Card for DFNDR-12M-BFloat16 This dataset contains synthetic captions, embeddings, and metadata for DFNDR-12M. The metadata has been generated using pretrained image-text models on DFN-12M, a uniformly sampled subset of 12.8M samples from DFN-2B. For details on how to use the metadata, please visit our ml-mobileclip repository. For code to generate multi-modal reinforced datasets at large scale see ml-mobileclip-dr repository. The float32 version of this dataset is… See the full description on the dataset page: https://huggingface.co/datasets/apple/DFNDR-12M-bf16.imagetext-to-image5 likes162 downloads5mo agoHugging Face07open-llm-leaderboard-old /details_one-man-army__UNA-34Beagles-32K-bf16-v1 Dataset Card for Evaluation run of one-man-army/UNA-34Beagles-32K-bf16-v1 Dataset automatically created during the evaluation run of model one-man-army/UNA-34Beagles-32K-bf16-v1 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_one-man-army__UNA-34Beagles-32K-bf16-v1.0 likes156 downloads3y agoHugging Face08open-llm-leaderboard-old /details_OpenBuddyEA__openbuddy-llama-30b-v7.1-bf16 Dataset Card for Evaluation run of OpenBuddyEA/openbuddy-llama-30b-v7.1-bf16 Dataset Summary Dataset automatically created during the evaluation run of model OpenBuddyEA/openbuddy-llama-30b-v7.1-bf16 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_OpenBuddyEA__openbuddy-llama-30b-v7.1-bf16.0 likes152 downloads3y agoHugging Face09opherlie /lora-test-case-NVIDIA-Nemotron-3-Super-120B-A12B-BF160 likes151 downloads6mo agoHugging Face10open-llm-leaderboard-old /details_OpenBuddy__openbuddy-codellama2-34b-v11.1-bf16 Dataset Card for Evaluation run of OpenBuddy/openbuddy-codellama2-34b-v11.1-bf16 Dataset Summary Dataset automatically created during the evaluation run of model OpenBuddy/openbuddy-codellama2-34b-v11.1-bf16 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_OpenBuddy__openbuddy-codellama2-34b-v11.1-bf16.0 likes138 downloads3y agoHugging Face11open-llm-leaderboard-old /details_Sao10K__Sensualize-Mixtral-bf16 Dataset Card for Evaluation run of Sao10K/Sensualize-Mixtral-bf16 Dataset automatically created during the evaluation run of model Sao10K/Sensualize-Mixtral-bf16 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Sao10K__Sensualize-Mixtral-bf16.0 likes117 downloads3y agoHugging Face12JWei05 /gemma4-e4b-rl100-hf-bf16-sdpa-topk128-overlay Gemma 4 E4B RL100 top-k-128 target overlay Precomputed off-policy distillation targets for the E4B-RL-step-100 to E2B experiment. Source traces: JWei05/gemma4-e4b-rl100-topk128-traces at revision 2b6e49a0a456ee9d67b16a1dc61785562bee90c9 Direction: Gemma 4 E4B RL step 100 teacher to Gemma 4 E2B base student Target engine: Hugging Face BF16 SDPA full forward Width: top-k 128 Stored target token IDs: int32 Stored target log-probabilities: float16 Causal alignment: response token… See the full description on the dataset page: https://huggingface.co/datasets/JWei05/gemma4-e4b-rl100-hf-bf16-sdpa-topk128-overlay.tabular10K<n<100K0 likes102 downloads2mo agoHugging Face13CatGoesMeow /PI0-BF16_100eps_Selection_Nuts_Bolt_Aug_10_inf-recordingThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 60, "total_frames": 77648, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:60" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/CatGoesMeow/PI0-BF16_100eps_Selection_Nuts_Bolt_Aug_10_inf-recording.tabularrobotics10K<n<100K0 likes102 downloads1mo agoHugging Face14LieUr /vivit-b16x2-k400-postblock5-bf16-activations ViViT-B/16x2 K400 post-block-5 bf16 activations This data-only repository contains the frozen 10,400-clip activation cache used by NeonByte L5: 3,200 train, 800 dev, and 6,400 eval activations. Each .safetensors file stores the 3,137 × 768 bfloat16 hidden state after ViViT block 5 in CLS|tubelet(t,y,x)-row-major order. The repository contains derived activations and provenance metadata only. It contains no raw video, frames, audio, model weights, executable scripts, or NeonByte… See the full description on the dataset page: https://huggingface.co/datasets/LieUr/vivit-b16x2-k400-postblock5-bf16-activations.1 likes96 downloads27d agoHugging Face15OALL /details_Kquant03__CognitiveFusion2-4x7B-BF16 Dataset Card for Evaluation run of Kquant03/CognitiveFusion2-4x7B-BF16 Dataset automatically created during the evaluation run of model Kquant03/CognitiveFusion2-4x7B-BF16. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Kquant03__CognitiveFusion2-4x7B-BF16.tabular100K<n<1M0 likes67 downloads2y agoHugging Face16nyu-dice-lab /lm-eval-results-Kquant03-Cognito-2x7B-bf16-private Dataset Card for Evaluation run of Kquant03/Cognito-2x7B-bf16 Dataset automatically created during the evaluation run of model Kquant03/Cognito-2x7B-bf16 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Kquant03-Cognito-2x7B-bf16-private.tabular100K<n<1M0 likes67 downloads2y agoHugging Face17open-llm-leaderboard-old /details_Edgerunners__yi-9b-may-ortho-baukit-13fail-3000total-bf160 likes66 downloads2y agoHugging Face18open-llm-leaderboard-old /details_Edgerunners__meta-llama-3-8b-instruct-hf-ortho-baukit-5fail-3000total-bf160 likes61 downloads2y agoHugging Face19nyu-dice-lab /lm-eval-results-Kquant03-Nanashi-2x7B-bf16-private Dataset Card for Evaluation run of Kquant03/Nanashi-2x7B-bf16 Dataset automatically created during the evaluation run of model Kquant03/Nanashi-2x7B-bf16 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Kquant03-Nanashi-2x7B-bf16-private.tabular100K<n<1M0 likes61 downloads2y agoHugging Face20open-llm-leaderboard-old /details_Weyaxi__MetaMath-una-cybertron-v2-bf16-Ties Dataset Card for Evaluation run of Weyaxi/MetaMath-una-cybertron-v2-bf16-Ties Dataset Summary Dataset automatically created during the evaluation run of model Weyaxi/MetaMath-una-cybertron-v2-bf16-Ties on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Weyaxi__MetaMath-una-cybertron-v2-bf16-Ties.0 likes60 downloads3y agoHugging Face21nyu-dice-lab /lm-eval-results-CultriX-NeuralTrix-bf16-private Dataset Card for Evaluation run of CultriX/NeuralTrix-bf16 Dataset automatically created during the evaluation run of model CultriX/NeuralTrix-bf16 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-CultriX-NeuralTrix-bf16-private.tabular100K<n<1M0 likes57 downloads2y agoHugging Face22aDaikiKamata /patch_policy_libero10_tipsv2_token_cache_bf16_evalvideon<1K0 likes56 downloads1d agoHugging Face23open-llm-leaderboard-old /details_OpenBuddy__openbuddy-llama-65b-v8-bf16 Dataset Card for Evaluation run of OpenBuddy/openbuddy-llama-65b-v8-bf16 Dataset Summary Dataset automatically created during the evaluation run of model OpenBuddy/openbuddy-llama-65b-v8-bf16 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_OpenBuddy__openbuddy-llama-65b-v8-bf16.0 likes55 downloads3y agoHugging Face24open-llm-leaderboard-old /details_pszemraj__pythia-31m-simplewiki-scratch-bf16 Dataset Card for Evaluation run of pszemraj/pythia-31m-simplewiki-scratch-bf16 Dataset Summary Dataset automatically created during the evaluation run of model pszemraj/pythia-31m-simplewiki-scratch-bf16 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_pszemraj__pythia-31m-simplewiki-scratch-bf16.0 likes55 downloads3y agoHugging Face25brandonmusic /GLM-5.3-BF16-shapleymcg-resume0 likes51 downloads1mo agoHugging Face26TAUR-dev /BF16kEval_FinEval_16k_fulleval__3args_ours-eval_rltabular10K<n<100K0 likes48 downloads11mo agoHugging Face27open-llm-leaderboard-old /details_grimjim__zephyr-wizard-kuno-royale-BF16-merge-7B Dataset Card for Evaluation run of grimjim/zephyr-wizard-kuno-royale-BF16-merge-7B Dataset automatically created during the evaluation run of model grimjim/zephyr-wizard-kuno-royale-BF16-merge-7B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_grimjim__zephyr-wizard-kuno-royale-BF16-merge-7B.0 likes45 downloads2y agoHugging Face28open-llm-leaderboard-old /details_Kquant03__Samlagast-7B-laser-bf16 Dataset Card for Evaluation run of Kquant03/Samlagast-7B-laser-bf16 Dataset automatically created during the evaluation run of model Kquant03/Samlagast-7B-laser-bf16 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Kquant03__Samlagast-7B-laser-bf16.0 likes44 downloads3y agoHugging Face29open-llm-leaderboard-old /details_dsvv-cair__alpaca-cleaned-llama-30b-bf16 Dataset Card for Evaluation run of dsvv-cair/alpaca-cleaned-llama-30b-bf16 Dataset Summary Dataset automatically created during the evaluation run of model dsvv-cair/alpaca-cleaned-llama-30b-bf16 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dsvv-cair__alpaca-cleaned-llama-30b-bf16.0 likes43 downloads3y agoHugging Face30open-llm-leaderboard-old /details_ConvexAI__Julianne-2x7B-bf16 Dataset Card for Evaluation run of ConvexAI/Julianne-2x7B-bf16 Dataset automatically created during the evaluation run of model ConvexAI/Julianne-2x7B-bf16 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ConvexAI__Julianne-2x7B-bf16.0 likes42 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.