r2ob12 hours agoRIS reduces self-attention complexity to $O(N \log N)$ using sparse stochastic geometry that fits within commodity memory limitshttps://www.nature.com/articles/s41598-026-59160-z
pestatije4 hours agoRIS-Kernel: A Model-Agnostic Architecture for Long-Context LLM Inference via Sparse Attention