SubQ – a sub-quadratic LLM built for multi-million token reasoning
SubQ is a sparse-attention AI model that skips unnecessary token calculations, delivering 56x faster performance than FlashAttention-2 at 12M tokens. It handles full codebases in a single prompt, matching GPT-5 and Claude Opus on key benchmarks.