Ax Timo Stoll, Chendi Qian, Ben Finkelshtein, Ali Parviz, Darius Weber, Fabrizio Frasca, Hadar Shavit, Antoine Siraudin, Arman Mielke, Marie Anastacio, Erik M\"uller, Maya Bechler-Speicher, Michael Bronstein, Mikhail Galkin, Holger Hoos, Mathias Niepert, Bryan Perozzi, Jan T\"onshoff, Christopher Morris 5/12/2026

GraphBench: Next-generation graph learning benchmarking

GraphBench provides standardized benchmarking suite for graph machine learning with consistent evaluation protocols for foundation models.

Ax Leyang Shen, Yang Zhang, Chun Kai Ling, Xiaoyan Zhao, Tat-Seng Chua 5/12/2026

CARL: Criticality-Aware Agentic Reinforcement Learning

CARL improves multi-step reinforcement learning for agents by identifying and optimizing criticality-aware action choices rather than treating all steps equally.

Ax Tiange Luo, Lajanugen Logeswaran, Jaekyeom Kim, Justin Johnson, Honglak Lee 5/12/2026

Selective LoRA for Visual Tokens and Attention Heads

Parameter-efficient fine-tuning method applying LoRA selectively to visual tokens and attention heads in vision-language models.

Ax Chenxiao Yu, Bowen Yi, Farzan Karimi-Malekabadi, Suhaib Abdurahman, Jinyi Ye, Shrikanth Narayanan, Yue Zhao, Morteza Dehghani 5/12/2026

Tracing Moral Foundations in Large Language Models

Study of how moral foundations are encoded and organized across 14 LLMs to determine if models have internal moral structure.

Ax Qiuyu Tian, Zequn Liu, Yiding Li, Fengyi Chen, Zequn Liu, Youyong Kong, Fan Guo, Yuyao Li, Jinjing Shen, Zhijing Xie, Yiyun Luo, Xin Zhang, Yingce Xia 5/12/2026

STAGE: A Full-Screenplay Benchmark for Reasoning over Evolving Storie

Benchmark of long-form narrative understanding using movie screenplays to evaluate story world reasoning and generation consistency.

Ax Qiyang Li, Sergey Levine 5/12/2026

Q-learning with Adjoint Matching

Novel reinforcement learning algorithm for continuous-action RL using adjoint matching with diffusion/flow-matching policies.

Ax Xavier Hu, Jinxiang Xia, Shengze Xu, Kangqi Song, Yishuo Yuan, Guibin Zhang, JinCheng Ren, Boyu Feng, Li Lu, Tieyong Zeng, Jiaheng Liu, Minghao Liu, He Zhu, Yuchen Eleanor Jiang, Wei Wang, Wangchunshu Zhou 5/12/2026

EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies

EcoGym benchmark for evaluating LLM-based agents on long-horizon planning and execution tasks in persistent interactive economic environments.