sparse
splade-bert-tiny-nqopensearch-neural-sparse-encoding-doc-v2-miniopensearch-neural-sparse-encoding-doc-v3-distillopensearch-neural-sparse-encoding-doc-v2-distillopensearch-neural-sparse-encoding-multilingual-v1splade-camembert-base-v2granite-embedding-30m-sparseopensearch-neural-sparse-encoding-v2-distill
Datasets
All datasets matching “sparse”Charge-040_0040-Sparse-MonoLanguage-Grounded_Sparse_Encoder_Training
Language-Grounded Sparse Encoder (LanSE) — Training Data
This repository hosts the AI-generated images and human annotation datasets accompanying the paper:
Human-like Content Analysis for Generative AI with Language-Grounded Sparse Encoders
Yiming Tang, Arash Lagzian, Srinivas Anumasa, Qiran Zou, Yingtao Zhu, Ye Zhang, Trang Nguyen, Yih-Chung Tham, Ehsan Adeli, Ching-Yu Cheng, Yilun Du, Dianbo Liu
National University of Singapore · Tsinghua University · Stanford University ·… See the full description on the dataset page: https://huggingface.co/datasets/DesmondYMTang2024/Language-Grounded_Sparse_Encoder_Training.Dr.Sparse-Granite42-8B-eval-b200-otf81-spgemm
Dr.Sparse — Granite 4.2 8B SpGEMM baseline (OTF-81)
ibm-granite/granite-4.2-8b 在 Dr.Sparse OTF 保留测试集上的 SpGEMM baseline。
levels 1-3(排除 level4),单轨迹无树搜索,每矩阵 10 轮迭代,B200 (sm_100)。
结果
level
矩阵
正确
跑赢 cuSPARSE
中位加速比
最大
level1_small
8
1
0
0.244
0.24
level2_medium
33
3
0
0.024
0.62
level3_large
38
4
0
0.092
1.00
合计
79
8
0
0.102
1.00
79 个矩阵里 8 个产出正确 kernel,无一跑赢 cuSPARSE。中位加速比 0.102
表示比 cuSPARSE 慢约十倍;最好的一个仅持平。
编译失败的错误类型分散:cudaMalloc 重载不匹配、const 限定符未去除、… See the full description on the dataset page: https://huggingface.co/datasets/DiogenesChen122/Dr.Sparse-Granite42-8B-eval-b200-otf81-spgemm.Dr.Sparse-Gemma4-12B-eval-b200-otf81-spgemm
Dr.Sparse — Gemma 4 12B base, SpGEMM on OTF-81
google/gemma-4-12B-it(未经微调)在 Dr.Sparse OTF 保留测试集上的 SpGEMM 评测。
levels 1-3(排除 level4),单轨迹无树搜索,每矩阵 10 轮迭代,B200 (sm_100)。
结果
level
矩阵
正确
跑赢 cuSPARSE
正确者中位加速比
最大
level1_small
8
0
0
-
-
level2_medium
31
0
0
-
-
level3_large
36
1
0
0.057
0.06
合计
75
1
0
0.057
0.06
81 个矩阵中 75 个跑完。其余 6 个因模型上下文超限(vLLM 返回 400)中止,未计入。
失败以真实编译错误为主,例如在 __shared__ 变量上写初始值、函数重复定义、
引用不存在的结构体成员。
对照
同一数据上的 SFT 版本见… See the full description on the dataset page: https://huggingface.co/datasets/DiogenesChen122/Dr.Sparse-Gemma4-12B-eval-b200-otf81-spgemm.Dr.Sparse-Gemma4-12B-SFT-eval-b200-otf81-spgemm
Dr.Sparse — Gemma 4 12B SFT, SpGEMM on OTF-81
Gemma 4 12B 经 Luna-10 v4b 数据 LoRA SFT 后,在 Dr.Sparse OTF 保留测试集上的 SpGEMM 评测。
levels 1-3(排除 level4),单轨迹无树搜索,每矩阵 10 轮迭代,B200 (sm_100)。
模型权重:DiogenesChen122/Dr.Sparse-Gemma4-12B-SFT-luna10-v4b
结果
level
矩阵
正确
跑赢 cuSPARSE
正确者中位加速比
最大
level1_small
8
3
0
0.263
0.49
level2_medium
34
6
2
0.373
2.15
level3_large
37
3
0
0.221
0.88
合计
79
12
2
0.314
2.15
81 个矩阵中 79 个跑完。其余 2 个因模型上下文超限(vLLM 返回 400)中止,未计入。… See the full description on the dataset page: https://huggingface.co/datasets/DiogenesChen122/Dr.Sparse-Gemma4-12B-SFT-eval-b200-otf81-spgemm.SparseVideoNav
SparseVideoNav Datasets
This repository contains the real-world navigation datasets released with OpenDriveLab/SparseVideoNav:
BVN: Beyond-the-View Navigation.
IFN: Instruction-Following Navigation.
Project links:
Project page: https://opendrivelab.com/SparseVideoNav
GitHub: https://github.com/OpenDriveLab/SparseVideoNav
Paper: https://arxiv.org/abs/2602.05827
Dataset Summary
SparseVideoNav studies real-world vision-language navigation with sparse future… See the full description on the dataset page: https://huggingface.co/datasets/OpenDriveLab/SparseVideoNav.
