Introducing SubQ - a major breakthrough in LLM intelligence. It is the first model built on a fully sub-quadratic sparse-attention architecture (SSA), And the first frontier model with a 12 million token context window which is: - 52x faster than FlashAttention at 1MM tokens -…↗
·@alex_whedon·May 5, 2026·model-release·context-window·flashattention·llm·sub-quadratic-sparse-attention