Ax Jatin Prakash, Aahlad Puli, Rajesh Ranganath 7/14/2026

Controllably Efficient Language Models

Framework enabling transformers to trade off inference efficiency and quality dynamically, controlling sparse/linear attention and convolutions per layer.

Ax Zhuoyun Du, Runze Wang, Huiyu Bai, Zouying Cao, Xiaoyong Zhu, Yu Cheng, Bo Zheng, Wei Chen, Haochao Ying 7/14/2026

Enabling Agents to Communicate Entirely in Latent Space

Interlat: enables LLM-based agents to communicate directly in latent space, bypassing discrete tokens for richer information exchange in collaborative problem-solving.

Ax Noam Koren, Rafael Moschopoulos, Kira Radinsky, Elad Hazan 7/14/2026

SFO: Learning PDE Operators via Spectral Filtering

Spectral Filtering Operator (SFO): neural operator for solving PDEs using Universal Spectral Basis to capture long-range nonlocal interactions.

Ax Nadav Benedek, Tomer Koren, Ohad Fried 7/14/2026

Gefen: Optimized Stochastic Optimizer

Gefen: memory-efficient optimizer reducing AdamW footprint by ~8x through moment state sharing and learned quantization for large-scale pretraining.

Ax Yunhe Li, Hao Shi, Wenhao Liu, Mengzhe Ruan, Hanxu Hou, Zhongxiang Dai, Shuang Qiu, Linqi Song 7/14/2026

DemoPSD: Disagreement-Modulated Policy Self-Distillation

DemoPSD: self-distillation method for training LLMs to reason by modulating policy disagreement, reducing overfitting and improving generalization.