Ax Sibo Zhu, Wenyi Wu, Kun Zhou, Stephen Wang, Biwei Huang 3/12/2026

Hybrid Self-evolving Structured Memory for GUI Agents

GUI agents using vision-language models with hybrid self-evolving structured memory to handle long-horizon workflows and diverse interfaces.

Ax Wenjing Zhang, Jiangze Yan, Jieyun Huang, Yi Shen, Shuming Shi, Ping Chen, Ning Wang, Zhaoxiang Liu, Kai Wang, Shiguo Lian 3/12/2026

HEAL: Hindsight Entropy-Assisted Learning for Reasoning Distillation

HEAL distills reasoning from large language models to smaller models via hindsight entropy-assisted learning, overcoming teacher limitations.

Ax Chuan Guo (Michael Pokorny), Juan Felipe Ceron Uribe (Michael Pokorny), Sicheng Zhu (Michael Pokorny), Christopher A. Choquette-Choo (Michael Pokorny), Steph Lin (Michael Pokorny), Nikhil Kandpal (Michael Pokorny), Milad Nasr (Michael Pokorny), Rai (Michael Pokorny), Sam Toyer, Miles Wang, Yaodong Yu, Alex Beutel, Kai Xiao 3/12/2026

IH-Challenge: A Training Dataset to Improve Instruction Hierarchy on Frontier LLMs

IH-Challenge dataset for training instruction hierarchy in LLMs. Addresses jailbreaks and prompt injection by establishing trust-ordered instruction conflict resolution.

Ax Baichuan Mo, Hanyong Xu, Ruoyun Ma, Jung-Hoon Cho, Dingyi Zhuang, Xiaotong Guo, Jinhua Zhao 3/12/2026

Large Language Models for Travel Behavior Prediction

Applies large language models for travel behavior prediction through natural language reasoning frameworks as alternative to numerical models.

Ax Raphi Kang, Yue Song, Georgia Gkioxari, Pietro Perona 3/12/2026

Is CLIP ideal? No. Can we fix it? Yes!

Addresses fundamental geometric limitations of CLIP's multimodal latent space for handling complex visual-textual interactions through geometric improvements.

Ax Morris Yau, Sharut Gupta, Valerie Engelmayer, Kazuki Irie, Stefanie Jegelka, Jacob Andreas 3/12/2026

Sequential-Parallel Duality in Prefix Scannable Models

Characterizes neural sequence models supporting sequential-parallel duality, analyzing architectures like Gated Linear Attention and Mamba.

Ax Kiril Bangachev, Guy Bresler, Iliyas Noman, Yury Polyanskiy 3/12/2026

Global Minimizers of Sigmoid Contrastive Loss

Theoretical analysis of sigmoid contrastive loss in CLIP-style models, explaining advantages of temperature and bias in SigLIP and SigLIP2.

Ax Bilge Acun, Prasoon Sinha, Newsha Ardalani, Sangmin Bae, Alicia Golden, Chien-Yu Lin, Meghana Madhyastha, Fei Sun, Neeraja J. Yadwadkar, Carole-Jean Wu 3/12/2026

Composer: A Search Framework for Hybrid Neural Architecture Design

Search framework for automatically discovering hybrid neural architectures combining attention, MLP, and other primitives beyond standard transformers.

Ax Naman Agarwal, Siddhartha R. Dalal, Vishal Misra 3/12/2026

Geometric Scaling of Bayesian Inference in LLMs

Investigates geometric substrate of Bayesian inference in production-scale language models (Pythia, Phi-2, Llama-3, Mistral).