SparseWake/sparsewake
SparseWake SparseWake is a synthetic benchmark for sparse temporal hydrodynamic sensing. ICLR 2027 release The expanded release adds controlled multi-source mixtures and common-prior nearest-source tasks, with complete core data banks, reference checkpoints, a small review supplement, and reproduction code with a frozen wake-library input. Download release iclr2027-v1.0rc2 The version page lists the three archives, exact sizes, checksums, extraction instructions… See the full description on the dataset page: https://huggingface.co/datasets/SparseWake/sparsewake.
0275
1# Artifact Audit2 3Audit date: 2026-05-044 5## Scope6 7Artifact folder: SparseWake anonymous release tree.8 9The manuscript source was not edited for this artifact-preparation task.10 11## Checks12 13| Check | Status | Notes |14|---|---|---|15| No absolute local paths | Pass | Text scan found no drive prefixes or Unix home paths. |16| No usernames, institution names, machine names, or personal emails | Pass | Text scan found no known local identifiers. |17| No WakeSchool ABM source code or DNS solver files | Pass | Artifact contains processed HDF5 datasets, Python benchmark utilities, result CSVs, figures, tables, and documentation only. |18| No unrelated data | Pass | Included full processed benchmark HDF5 files, sample HDF5, result CSVs, configs, documentation, and generated verification outputs. |19| All included HDF5 files have checksums | Pass | See `data/checksums.sha256`. |20| `verify_dataset.py` runs on sample data | Pass | Verified `data/sample/sparsewake_sample.h5`. |21| Table-generation script runs from stored CSVs | Pass | `scripts/reproduce_tables.py` wrote CSV tables to `tables/`. |22| Main figure-generation script runs from stored CSVs | Pass | `scripts/make_main_figures.py` wrote PDF/SVG figures. |23| Supplement figure-generation script runs from stored CSVs | Pass | `scripts/make_supp_figures.py` wrote PDF/SVG figures. |24| Quick training smoke test runs on sample data | Pass | `scripts/train_temporal_mlp.py --quick` completed and wrote quick metrics. |25| README, dataset card, evaluation card, provenance, and Croissant metadata exist | Pass | Required documentation files are present. |26| Manuscript numbers traceable to CSV or result summary | Pass | Clean summaries are in `data/results/`; source traceability is documented in the cards and manifest. |27 28## Tested Commands29 30```bash31python scripts/verify_dataset.py --data data/sample/sparsewake_sample.h532python scripts/train_temporal_mlp.py --config configs/main_v04.yaml --data data/sample/sparsewake_sample.h5 --quick33python scripts/reproduce_tables.py --results data/results --out tables34python scripts/make_main_figures.py --results data/results --out figures35python scripts/make_supp_figures.py --results data/results --out figures36```37 38## Manual Items Before Upload39 40- Confirm `data/processed/download_manifest.json` uses relative paths or anonymous hosted paths.41- Run a formal Croissant validator if required by the final artifact platform.42- Confirm final license wording is CC BY 4.0.43 