Ax Ari Pakman, Lior Kreimer, Yakir Berchenko 7/1/2026

Revisiting the Volume Hypothesis

Investigation of volume hypothesis explaining neural network generalization through loss-landscape basin geometry and implicit bias of SGD.

Ax Zena Al-Khalili, Rafi Hakim, Dietrich Klakow, Ji-Ung Lee 7/1/2026

Fork-Think with Confidence

Decide-first-then-think method for improving LLM reasoning tasks by identifying high-confidence decision points before generation.

Ax Philippe Chlenski, Zachariah Carmichael, Ayush Warikoo, Chia-Tse Shao, Yingxiao Ye, Aobo Yang, Vivek Miglani, Nehal Bandi 7/1/2026

Surrogate Fidelity: When Can Open LLMs Explain Closed Ones?

Studies when open-source LLMs can explain closed models through mechanistic interpretability, evaluating surrogate fidelity across prediction and representation levels.

Ax Yuanda Xu, Zhengze Zhou, Hejian Sang, Xiaomin Li, Jiaxin Zhang, Xinchen Du, Zhipeng Wang, Alborz Geramifard 7/1/2026

TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning

TRIAGE proposes role-typed credit assignment for agentic RL, improving upon GRPO by differentiating credit across action types in agent-environment interactions.