[–] r2ob 28d ago ↗ RIS reduces self-attention complexity to $O(N \log N)$ using sparse stochastic geometry that fits within commodity memory limitshttps://www.nature.com/articles/s41598-026-59160-z
[–] pestatije 28d ago ↗ RIS-Kernel: A Model-Agnostic Architecture for Long-Context LLM Inference via Sparse Attention
2 comments
[ 0.23 ms ] story [ 10.4 ms ] threadhttps://www.nature.com/articles/s41598-026-59160-z