Ax Chika Maduabuchi 7/1/2026

Entropy-Controlled Flow Matching

Entropy-controlled flow matching constrains information geometry in generative models to prevent semantic mode collapse.

Ax Huaiyang Wang, Xiaojie Li, Xiaohan Wang, Zhixia Zhang, Xiaodong Lu, Zixuan Huang, Jiajun Chai, Guojun Yin, Deqing Wang, Haoyi Zhou, Yaodong Yang, Jianxin Li, Yikun Ban 7/1/2026

Policy Improvement Reinforcement Learning

RL post-training method for LLMs/agents that explicitly verifies policy improvements over predecessors before updating.

Ax Siyu Chen, Miao Lu, Beining Wu, Heejune Sheen, Fengzhuo Zhang, Shuangning Li, Zhiyuan Li, Jose Blanchet, Tianhao Wang, Zhuoran Yang 7/1/2026

INFUSER: Influence-Guided Self-Evolution Improves Reasoning

INFUSER iterative co-training framework enables LLMs to self-improve reasoning with minimal external supervision using influence-guided generation.

Ax Lukas Fesser, Hanlin Zhang, Michelle M. Li, Eric Wang, Bryan Perozzi, Shekoofeh Azizi, Sham M. Kakade, Marinka Zitnik 7/1/2026

How Post-Training Shapes Biological Reasoning Models

Studies how post-training stages in biological reasoning models affect performance and generalization across genomics, transcriptomics, and proteins.

Ax Soham Bhattacharjee, Dushyant Singh Chauhan, Salem Lahlou, Martin Takac, Nils Lukas 7/1/2026

Entropy-Gated Latent Recursion

Inference-time scaling method using layer-span recursion and entropy gating to improve language model reasoning without token sampling alone.

Ax Yijie Jin, Jiajun Xu, Yuxuan Liu, Chenkai Xu, Yi Tu, Jiajun Li, Dandan Tu, Xiaohui Yan, Kai Yu, Pengfei Liu, Zhijie Deng 7/1/2026

Multi-Block Diffusion Language Models

Multi-block diffusion language models enabling concurrent decoding of consecutive blocks for inter-block parallelism and flexible-length generation.

Ax Eduardo Sebasti\'an, Nicolas Pfitzer, Ajay Shankar, Amanda Prorok 7/1/2026

Prompting Robot Teams with Natural Language

Framework for prompting multi-robot teams with natural language, decomposing collaborative tasks without runtime LLM calls.