Transformers and self-attention rewrote the rules of AI, but they are starting to hit hard limits on speed, memory, and context. A new wave of architectures—state space models like Mamba, convolutional operators like Hyena, and hybrid RNNs like RWKV—are quietly answering the question "what comes after attention" with very different ideas about how sequence models should work.