Ax Dengjia Zhang, Alexander Martin, William Jurayj, Kenton Murray, Benjamin Van Durme, Reno Kriz 4/13/2026

Unified Multimodal Uncertain Inference

Multimodal inference task with text, audio, video for producing calibrated probability estimates of hypotheses with fine-grained uncertainty.

Ax Jinghan Zhang, Fengran Mo, Tharindu Cyril Weerasooriya, Ruimin Dai, Xiaoyan Han, Yanjie Fu, Dakuo Wang, Kunpeng Liu 4/13/2026

StaRPO: Stability-Augmented Reinforcement Policy Optimization

RL framework for improving LLM reasoning by optimizing for logical consistency and structural integrity of reasoning processes, not just final answers.

Ax Stefan Andreas Baumann, Jannik Wiese, Tommaso Martorella, Mahdi M. Kalayeh, Bj\"orn Ommer 4/13/2026

Envisioning the Future, One Step at a Time

Video prediction model representing scene dynamics as sparse point trajectories for efficient future frame synthesis.

Ax ShengYun Peng, Eric Smith, Ivan Evtimov, Song Jiang, Pin-Yu Chen, Hongyuan Zhan, Haozhu Wang, Duen Horng Chau, Mahesh Pasupuleti, Jianfeng Chi 4/13/2026

Large Reasoning Models Learn Better Alignment from Flawed Thinking

RECAP: RL method for safety alignment in large reasoning models, teaching critical evaluation of flawed premises via counter-aligned prefilling.