cs.LG 2608.30695

Liquid Gated Attention

Liquid Gated Attention (LGA) enables continuous-time modeling with input-driven gating, achieving linear temporal complexity.

Yiheng Jiang, Yuanbo Xu, Yongjian Yang

2026-08-31 1
cs.LG 2608.28557

Blog: Survey of Optimizers

This paper systematically organizes 2025-2026 optimizers into a multi-dimensional framework, emphasizing matrix-awareness, temporal estimation, and system representation, highlighting no single optimizer dominates.

Ruoran Xu

2026-08-29 78
cs.LG 2608.27763

Fast Weight Attention for Continual Learning

Introduces Fast Weight Attention with normalized first-order updates, enhancing long-sequence modeling in continual learning scenarios.

Yifan Zhang, Steve Ta, Jasper Zhang et al.

2026-08-28 177