Ax Iv\'an Arcuschin, Jett Janiak, Robert Krzyzanowski, Senthooran Rajamanoharan, Neel Nanda, Arthur Conmy 6/1/2026

Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

Research showing chain-of-thought reasoning in LLMs exhibits unfaithfulness on natural prompts without explicit biases, revealing reasoning discrepancies.

Ax Subham Sekhar Sahoo, Zhihan Yang, Yash Akhauri, Johnna Liu, Deepansha Singh, Zhoujun Cheng, Zhengzhong Liu, Eric Xing, John Thickstun, Arash Vahdat 6/1/2026

Esoteric Language Models: A Family of Any-Order Diffusion LLMs

Eso-LMs: Family of diffusion language models interpolating autoregressive and masked diffusion paradigms with parallel generation and KV caching.

Ax Facheng Yu, Ronak Mehta, Alex Luedtke, Zaid Harchaoui 6/1/2026

Stochastic Gradients under Nuisances

Theoretical analysis of stochastic gradient optimization with unknown nuisance parameters, establishing convergence guarantees.

Ax Axel Mezini, Elena Umili, Ivan Donadello, Fabrizio Maria Maggi, Matteo Mancanelli, Fabio Patrizi 6/1/2026

Neuro-Symbolic Predictive Process Monitoring

Neuro-symbolic approach integrating deep learning with temporal logic for business process monitoring and suffix prediction.

Ax Ethan Shen, Daniel Tormoen, Saurabh Shah, Ali Farhadi, Tim Dettmers 6/1/2026

SERA: Soft-Verified Efficient Repository Agents

SERA framework for efficiently training open-weight repository-specific coding agents with soft verification without expensive full training.

Ax Yannis Montreuil, Le\"ina Montreuil, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi 6/1/2026

Learning-to-Defer with Expert-Conditional Advice

Learning-to-defer framework allowing dynamic selection of expert-specific information like retrieved documents and tool outputs.

Ax Yunhe Li, Hao Shi, Bowen Deng, Wei Wang, Mengzhe Ruan, Hanxu Hou, Zhongxiang Dai, Siyang Gao, Chao Wang, Shuang Qiu, Linqi Song 6/1/2026

Learning to Reason with Insight for Informal Theorem Proving

DeepInsight method for informal theorem proving with LLMs by identifying core solution techniques through insight recognition.

Ax Rajinder Sandhu, Di Mu, Cheng Chang, Md Shahriar Tasjid, Himanshu Rai, Maksims Volkovs, Ga Wu 6/1/2026

Aligning Dense Retrievers with LLM Utility via Distillation

Utility-Aligned Embeddings framework combining dense retrieval with LLM utility for improved RAG performance and computational efficiency.

Ax Dang Hoang Duy, Yannis Montreuil, Maxime Meyer, Axel Carlier, Lai Xing Ng, Wei Tsang Ooi 6/1/2026

Online Learning-to-Defer with Varying Experts

Online learning-to-defer algorithm for routing queries between models and varying expert pools with bandit feedback.

Ax Benjamin Schneider, Xavier Schneider, Victor Zhong, Sun Sun 6/1/2026

ASH: Agents that Self-Hone via Embodied Learning

ASH: agentic system for embodied policy learning from unlabeled internet video using self-improvement loop with inverse dynamics models.