Ax John Kirchenbauer, Brian R. Bartoldson, Bhavya Kailkhura, Tom Goldstein 7/2/2026

Watermarking for Proprietary Dataset Protection

Uses output watermarking techniques to make training data membership inference more tractable for generative language models.

Ax Anindya Sarkar, Nasik Muhammad Nafi, Isaac Lyngaas, Muralikrishnan Gopalakrishnan Meena, Yevgeniy Vorobeychik 7/2/2026

PAPA: Online Personalized Active Preference Alignment

Diffusion models fine-tuned via RL for personalized recommender systems using interactive user feedback.

Ax Dan Ley, Giang Nguyen, Himabindu Lakkaraju, Julius Adebayo 7/2/2026

Prototype Language Models

arXiv paper introducing prototype language models that make training data influence explicit and traceable for interpretability.

Ax Christopher Lindenberg, Kashyap Chitta 7/2/2026

Valdi: Value Diffusion World Models

arXiv paper on Value Diffusion World Models combining diffusion models with model predictive control for latent planning.

Ax Ludwig Winkler, Andrew Leaver-Fay, Joseph Kleinhenz, Pan Kessel 7/2/2026

Diffeomorphic Optimization

arXiv paper on diffeomorphic optimization for solving objectives on low-dimensional manifolds using diffusion models.

Ax Jingwei Song, Haofeng Xu, Jie Xiao, Chengke Bao, Jingwei Shi, Pengbin Feng, Weixun Wang, Yuhang Han, Chuan Wu, Linfeng Zhang, Bill Shi 7/2/2026

Staleness-Learning Rate Scaling Laws for Asynchronous RLHF

arXiv paper studying staleness effects in asynchronous RLHF systems with scaling laws for policy optimization.

Ax Hao Huang 7/2/2026

Muon as a Residual Connection

arXiv paper proposing Muon optimizer as implicit residual connection during neural network training, explaining its effectiveness.