The paper reports that its SSD implementation is 2–8× faster than Mamba's fused associative scan for large state expansion, is faster than FlashAttention-2 at sequence lengths of 2K and above, and is 6× faster than FlashAttention-2 at sequence length 16K. At sequence length 4K, increasing state expansion slows the optimized Mamba scan approximately linearly, whereas SSD shows little slowdown. · CiteArk
The paper reports that its SSD implementation is 2–8× faster than Mamba's fused associative scan for large state expansion, is faster than FlashAttention-2 at sequence lengths of 2K and above, and is 6× faster than FlashAttention-2 at sequence length 16K. At sequence length 4K, increasing state expansion slows the optimized Mamba scan approximately linearly, whereas SSD shows little slowdown.
Source: paper_markdown:Figure 10 and Section 9.3, PDF pages 29–31; introduction, PDF pages 2–3
0 supported0 disputed or conflicting0 not assessed
The complete result group stays with its claim. Assessed rows come first; their judgments do not replace the whole-claim conclusion. Expand a row for conditions, sources, and evidence history.
Each row is a measurement with its reported value, latest available observation, and cumulative evidence state.