avithal/repro-sf-mamba
Reproduction: SF-Mamba: Rethinking State Space Model for Vision Independent reproduction of ICML 2026 paper R9AUrEgZEq (arXiv 2603.16423) by Yoshimura, Hayashi, Hoshino, Wang, Ohashi (Sony). No official code or checkpoints are released ("We will release the source code after publication"). SF-Mamba builds on the public NVlabs/MambaVision backbone (MambaVisionMixer(d_state=8, d_conv=3, expand=1), SSM on dim/2 channels). What is reproduced Claim Approach… See the full description on the dataset page: https://huggingface.co/datasets/avithal/repro-sf-mamba.
Reproduction: SF-Mamba: Rethinking State Space Model for Vision
Independent reproduction of ICML 2026 paper R9AUrEgZEq (arXiv 2603.16423) by Yoshimura, Hayashi, Hoshino, Wang, Ohashi (Sony).
No official code or checkpoints are released ("We will release the source code after publication"). SF-Mamba builds on the public NVlabs/MambaVision backbone (MambaVisionMixer(d_state=8, d_conv=3, expand=1), SSM on dim/2 channels).
What is reproduced
Rerun
# CPU-ok correctness parts
python claim2_bench.py --out claim2.json # fp64 fold+reset equivalence
python claim1_swap.py --steps 1500 --seeds 3 # mechanism test (GPU faster)
# A100 job (HF Jobs)
hf jobs run --flavor a100-large --timeout 3600 -v ./:/repro \
-v hf://buckets/<user>/<bucket>:/data \
pytorch/pytorch:2.4.0-cuda12.1-cudnn9-devel bash /repro/job.sh