SparseDitto: An Agentic Sparse Compilation Framework through Architecture-Aware Synthesis on GPUs
Updated 3 h ago · first seen 11 Sept 2026
paper_01M294FTDN68YP1YNT916V3AH0
- Published
- 11 Sept 2026
- T1 · 3 h ago
- arXiv
- 2608.05033
- T1 · 3 h ago
- Category
- cs.DC
- T1 · 3 h ago
Abstract
-cross Abstract: Sparse matrix computation performance on GPU depends on how representation and execution schedule match the input structure and target hardware. No single implementation consistently dominates across sparsity patterns, operators, and hardwares. Existing sparse compilers and specialized systems cannot cover all of them simultaneously. We present SparseDitto, an agentic sparse compilation framework for sparse matrix computation on GPUs. It jointly synthesizes representation, execution schedule, and hardware mapping in a unified compilation plan. Structural analysis and a learned template-ranking prior guide architecture-aware synthesis. LLM-guided lowering realizes each plan as CUDA code, while target-GPU profiling drives plan refinement. SparseDitto covers multiple operators, e.g., SpMV, SpMM, and SpGEMM, and various representations within one framework. It can also automatically adapt to different hardwares. Across various SuiteSparse matrices, SparseDitto achieves geometric-mean speedups over cuSPARSE of $2.68\times$ on an NVIDIA RTX PRO 6000 and $2.79\times$ on an NVIDIA H200 (up to 146.61$\times$). Its generated SpMM kernels accelerate full-batch GCN training by up to $3.39\times$.
Authors 6
Shiyang Li, Guangyan Sun, Jinwei Tang, Yanzhi Wang, Mingyi Hong, Caiwen Ding
Specification
- Official page
Source:arXiv (Atom API + RSS)T1observed 3 h agohigh
- Arxiv announce type
- replace
Source:arXiv (Atom API + RSS)T1observed 3 h agohigh
- arXiv id
- 2608.05033
Source:arXiv (Atom API + RSS)T1observed 3 h agohigh
- Categories
- cs.DC, cs.LG
Source:arXiv (Atom API + RSS)T1observed 3 h agohigh
Source:arXiv (Atom API + RSS)T1observed 3 h agohigh
- Primary category
- cs.DC
Source:arXiv (Atom API + RSS)T1observed 3 h agohigh
- Published
- 11 Sept 2026
Source:arXiv (Atom API + RSS)T1observed 3 h agohigh
Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →
Provenance
Attributed facts
9
Source tiers
T19
Freshest observation
3 h ago
Conflicts
None
No models linked to this paper yet.
- Authors
- Shiyang Li, Guangyan Sun, Jinwei Tang
As of
Rewind the record: see this entity's attributes exactly as AI Atlas knew them on a given day.
Claim history · PDF
PDFpdf_url1
| Value | Valid from → to | Status | Source | Confidence | Extractor |
|---|---|---|---|---|---|
| https://arxiv.org/pdf/2608.05033 | → current | current | arXiv (Atom API + RSS)T1 | high | deterministic |
Claims are temporal and append-only: a new observation closes the previous claim (valid_to) instead of overwriting it. Conflicting claims from different sources are kept side by side and flagged — never averaged. Methodology →
- New paperPaperSparseDitto: An Agentic Sparse Compilation Framework through Architecture-Aware Synthesis on GPUs
New paper: SparseDitto: An Agentic Sparse Compilation Framework through Architecture-Aware Synthesis on GPUs
arxiv
| Source | Document | Type | Tier | Last observed | Snapshots |
|---|---|---|---|---|---|
| arXiv (Atom API + RSS) | rss.arxiv.org/rss/cs.LG | feed | T1· Official | 1 h ago | 1 |
Tier 1 = official/primary, 2 = quality secondary, 3 = community, 4 = unverified. Every snapshot is archived; see all sources and the methodology.