Ax Ali Behrouz, Zeman Li, Yuan Deng, Peilin Zhong, Meisam Razaviyayn, Vahab Mirrokni 3/2/2026

Memory Caching: RNNs with Growing Memory

Memory caching architecture enabling RNNs with growing memory capacity and subquadratic complexity as alternative to Transformers for sequence modeling.

Ax Weinan Dai, Hanlin Wu, Qiying Yu, Huan-ang Gao, Jiahao Li, Chengquan Jiang, Weiqiang Lou, Yufan Song, Hongli Yu, Jiaze Chen, Wei-Ying Ma, Ya-Qin Zhang, Jingjing Liu, Mingxuan Wang, Xin Liu, Hao Zhou 3/2/2026

CUDA Agent: Large-Scale Agentic RL for High-Performance CUDA Kernel Generation

Agentic RL system using LLMs for high-performance CUDA kernel generation at scale, outcompeting traditional compiler-based approaches.

Ax Masahiro Kato 3/2/2026

General Bayesian Policy Learning

General Bayes framework for policy learning where decision rules are the target rather than outcome prediction.

Ax Nathanael Jo, Nikhil Garg, Manish Raghavan 3/2/2026

The Subjectivity of Monoculture

Analysis of monoculture in LLMs showing agreement metrics depend on subjective baseline assumptions for independence.

Ax Shengqu Cai, Weili Nie, Chao Liu, Julius Berner, Lvmin Zhang, Nanye Ma, Hansheng Chen, Maneesh Agrawala, Leonidas Guibas, Gordon Wetzstein, Arash Vahdat 3/2/2026

Mode Seeking meets Mean Seeking for Fast Long Video Generation

Proposes training paradigm decoupling local fidelity from long-term coherence for scaling video generation from seconds to minutes.