Team Ai
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Evan-Lin /metric-mamba-ml2021-hungyi-corpus Dataset Card for "metric-mamba-ml2021-hungyi-corpus" More Information needed audio10K<n<100K0 likes473 downloads2y agoHugging Face02rookierufus /Vjepa_mamba_datasettabular10K<n<100K0 likes187 downloads4mo agoHugging Face03open-llm-leaderboard-old /details_CobraMamba__mamba-gpt-3b-v3 Dataset Card for Evaluation run of CobraMamba/mamba-gpt-3b-v3 Dataset Summary Dataset automatically created during the evaluation run of model CobraMamba/mamba-gpt-3b-v3 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CobraMamba__mamba-gpt-3b-v3.0 likes158 downloads3y agoHugging Face04Evan-Lin /mamba-ml2021-hungyi-corpus Dataset Card for "mamba-ml2021-hungyi-corpus" More Information needed audio10K<n<100K0 likes126 downloads2y agoHugging Face05open-llm-leaderboard-old /details_CobraMamba__mamba-gpt-3b Dataset Card for Evaluation run of CobraMamba/mamba-gpt-3b Dataset Summary Dataset automatically created during the evaluation run of model CobraMamba/mamba-gpt-3b on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CobraMamba__mamba-gpt-3b.0 likes112 downloads3y agoHugging Face06open-llm-leaderboard-old /details_TRI-ML__mamba-7b-rw Dataset Card for Evaluation run of TRI-ML/mamba-7b-rw Dataset automatically created during the evaluation run of model TRI-ML/mamba-7b-rw on the Open LLM Leaderboard. The dataset is composed of 62 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TRI-ML__mamba-7b-rw.0 likes108 downloads2y agoHugging Face07adimnaku /fpga_cost_model_kernel_data_mamba_p2 FPGA HLS Kernel Cost-Model Data Evolved Vitis HLS C++ kernels paired with their ground-truth Vitis HLS csynth results. Each row is one generated program from an evolutionary FPGA optimisation run, linked to its kernel source, evaluator report.json, and raw synthesis report. Each row carries a split label: train marks the original benchmarks used to fit the analytical cost model's learned correction term, and holdout marks benchmarks added afterwards that were not used for… See the full description on the dataset page: https://huggingface.co/datasets/adimnaku/fpga_cost_model_kernel_data_mamba_p2.tabulartabular-regressionn<1K0 likes87 downloads3mo agoHugging Face08Evan-Lin /snr-mamba-ml2021-hungyi-corpus Dataset Card for "snr-mamba-ml2021-hungyi-corpus" More Information needed audio10K<n<100K0 likes71 downloads2y agoHugging Face09open-llm-leaderboard-old /details_CobraMamba__mamba-gpt-3b-v2 Dataset Card for Evaluation run of CobraMamba/mamba-gpt-3b-v2 Dataset Summary Dataset automatically created during the evaluation run of model CobraMamba/mamba-gpt-3b-v2 on the Open LLM Leaderboard. The dataset is composed of 61 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CobraMamba__mamba-gpt-3b-v2.0 likes59 downloads3y agoHugging Face10Srishti280992 /repro-mamba-icl-outliers-actual Actual Mamba ICL Outlier Reproduction Artifacts Measured reproduction artifacts for ICML 2026 paper C41aLahRXZ, "How Can Mamba Learn In Context with Outliers and Generalize Provably?". The scripts implement the paper synthetic setup: Eq. (3), Definitions 1-2, hinge-loss SGD, one-layer Mamba and linear Transformer control. Outputs in actual_outputs/ are from 5 seeds, 4000 SGD iterations, batch 128, and 4000 prompts per evaluation point. HF GPU evidence:… See the full description on the dataset page: https://huggingface.co/datasets/Srishti280992/repro-mamba-icl-outliers-actual.text0 likes54 downloads2mo agoHugging Face11MambaHub /InsectMambaData1 likes46 downloads2y agoHugging Face12introvoyz041 /U-Mamba20 likes45 downloads3mo agoHugging Face13open-llm-leaderboard-old /details_CobraMamba__mamba-gpt-7b-v1 Dataset Card for Evaluation run of CobraMamba/mamba-gpt-7b-v1 Dataset Summary Dataset automatically created during the evaluation run of model CobraMamba/mamba-gpt-7b-v1 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CobraMamba__mamba-gpt-7b-v1.0 likes42 downloads3y agoHugging Face14rspiocbis /Mamba2InLlama_0_875-best_of_n-PRM-3859129tabularn<1K0 likes41 downloads2y agoHugging Face15avithal /repro-mamba-icl-outliers Reproduction: How Can Mamba Learn In Context with Outliers and Generalize Provably? Independent reproduction of ICML 2026 paper C41aLahRXZ (arXiv 2510.00399, titled on arXiv "Can Mamba Learn In Context with Outliers? A Theoretical Generalization Analysis"), by Li, Lu, Cui, Chen, Wang. No official code was released. This is a from-scratch implementation of the paper's own theoretical model and synthetic ICL experiments. What is implemented mamba_icl.py — the… See the full description on the dataset page: https://huggingface.co/datasets/avithal/repro-mamba-icl-outliers.0 likes41 downloads3mo agoHugging Face16W4ng1204 /pretrain_dataset_mamba10K<n<100K0 likes38 downloads2y agoHugging Face17avithal /repro-sf-mamba Reproduction: SF-Mamba: Rethinking State Space Model for Vision Independent reproduction of ICML 2026 paper R9AUrEgZEq (arXiv 2603.16423) by Yoshimura, Hayashi, Hoshino, Wang, Ohashi (Sony). No official code or checkpoints are released ("We will release the source code after publication"). SF-Mamba builds on the public NVlabs/MambaVision backbone (MambaVisionMixer(d_state=8, d_conv=3, expand=1), SSM on dim/2 channels). What is reproduced Claim Approach… See the full description on the dataset page: https://huggingface.co/datasets/avithal/repro-sf-mamba.0 likes38 downloads3mo agoHugging Face18open-llm-leaderboard /tiiuae__falcon-mamba-7b-detailsgated Dataset Card for Evaluation run of tiiuae/falcon-mamba-7b Dataset automatically created during the evaluation run of model tiiuae/falcon-mamba-7b The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__falcon-mamba-7b-details.tabular10K<n<100K1 likes37 downloads2y agoHugging Face19MambaRetriever /SPScanner Dataset We release the training and evaluation dataset of Single-Pass Scanner. Our train set is mambaretriever_train.jsonl, our test set by categories is mambaretriever_test_per_category.json, and out test set by dataset is mambaretriever_test.json. For more information about Single-Pass Scanner and the details of the datasets, check the Single-Pass Scanner Github question-answering2 likes37 downloads2y agoHugging Face20open-llm-leaderboard /tiiuae__Falcon3-Mamba-7B-Base-detailsgated Dataset Card for Evaluation run of tiiuae/Falcon3-Mamba-7B-Base Dataset automatically created during the evaluation run of model tiiuae/Falcon3-Mamba-7B-Base The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__Falcon3-Mamba-7B-Base-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face21rspiocbis /MambaInLlama_0_50-best_of_n-PRM-3517996tabular1K<n<10K0 likes35 downloads2y agoHugging Face22open-llm-leaderboard-old /details_CobraMamba__mamba-gpt-7b-v2 Dataset Card for Evaluation run of CobraMamba/mamba-gpt-7b-v2 Dataset Summary Dataset automatically created during the evaluation run of model CobraMamba/mamba-gpt-7b-v2 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CobraMamba__mamba-gpt-7b-v2.0 likes32 downloads3y agoHugging Face23mamba413 /train_data_imdb_simu HH-RLHF-Helpful-Base Dataset Summary The HH-RLHF-Helpful-Base dataset is a processed version of Anthropic's HH-RLHF dataset, specifically curated to train models using the TRL library for preference learning and alignment tasks. It contains pairs of text samples, each labeled as either "chosen" or "rejected," based on human preferences regarding the helpfulness of the responses. This dataset enables models to learn human preferences in generating helpful responses… See the full description on the dataset page: https://huggingface.co/datasets/mamba413/train_data_imdb_simu.tabular10K<n<100K0 likes31 downloads2y agoHugging Face242796gauravc /mamba-30m-dataset0 likes30 downloads9mo agoHugging Face25rspiocbis /Mamba2InLlama_0_875-best_of_n-PRM-3757864tabularn<1K0 likes28 downloads2y agoHugging Face26Klahadore /ptm_mamba_datasettabular100K<n<1M1 likes27 downloads1y agoHugging Face27open-llm-leaderboard /tiiuae__Falcon3-Mamba-7B-Instruct-detailsgated Dataset Card for Evaluation run of tiiuae/Falcon3-Mamba-7B-Instruct Dataset automatically created during the evaluation run of model tiiuae/Falcon3-Mamba-7B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__Falcon3-Mamba-7B-Instruct-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face28rspiocbis /Mamba2InLlama_0_50-best_of_n-PRM-3619944tabularn<1K0 likes25 downloads2y agoHugging Face29open-llm-leaderboard-old /details_CobraMamba__mamba-gpt-3b-v4 Dataset Card for Evaluation run of CobraMamba/mamba-gpt-3b-v4 Dataset Summary Dataset automatically created during the evaluation run of model CobraMamba/mamba-gpt-3b-v4 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CobraMamba__mamba-gpt-3b-v4.0 likes24 downloads3y agoHugging Face30ARotting /micro-mamba-memory MicroMamba MicroMamba is a small-compute test of input-dependent state-space dynamics. Each sequence contains distracting symbols, a few marked symbols, and a final query asking for one marked item by ordinal position. Solving the task requires selective storage and retrieval rather than ordinary next-token statistics. The model uses a compact Mamba-inspired block with: a causal depthwise convolution; learned stable diagonal state dynamics; input-dependent discretization, input… See the full description on the dataset page: https://huggingface.co/datasets/ARotting/micro-mamba-memory.0 likes23 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.