#Sparse-attention
Showing 6 of 6 repositories tagged #sparse-attention, ranked by stars
thu-ml
SpargeAttn
[ICML2025] SpargeAttention: A training-free sparse attention that accelerates any model inference.
Score
100
โ
1.0k
โ 99
+1/day
Cuda
lucidrains
native-sparse-attention-pytorch
Implementation of the sparse attention pattern proposed by the Deepseek team in their "Native Sparse Attention" paper
Score
0
โ
811
โ 53
โ
Python
HKUSTDial
flash-sparse-attention
Trainable fast and memory-efficient sparse attention
Score
0
โ
743
โ 53
โ
Python
AarambhDevHub
aarambh-studio
๐ฆ Decoder-only LLM built from scratch in pure Rust using Candle โ no Python, no PyTorch. Gated DeltaNet + sparse attention, fine-grained MoE, native video/document understanding, long-horizon tool agents, quantization-aware training. Scales: Tiny (25M) to Large (1.3B).
Score
67
โ
74
โ 20
+11/day
Rust
skylight-org
sparse-attention-hub
Advancing the frontier of efficient AI
Score
0
โ
68
โ 13
โ
Python
Infini-AI-Lab
vortex_torch
Vortex: Programmable Sparse Attention for Agents as Algorithm Designers
Score
33
โ
68
โ 12
โ
Python