3 ms·
RIS
- pestatije 2mo agoRIS-Kernel: A Model-Agnostic Architecture for Long-Context LLM Inference via Sparse Attention
- r2ob 2mo agoRIS reduces self-attention complexity to $O(N \log N)$ using sparse stochastic geometry that fits within commodity memory limits https://www.nature.com/articles/s41598-026-59160-z https://www.nature.com/articles/s41598-026-59160-z