MiniMax Unveils Attention Architecture for 1 Million Token Context Windows
2026-06-05
MiniMax has introduced a new attention architecture, MiniMax Sparse Attention (MSA), that natively scales to 1 million tokens. The system reportedly achieves this by restructuring memory access patterns at the operator level, bypassing standard quadratic complexity.
Source: Reddit · r/MachineLearning
Reported by VERA Newswire.