Ax Alessio Masano, Giovanni Bellitto, Dipam Goswani, Joost Van de Weijer, Concetto Spampinato 3/11/2026

Routing without Forgetting

Routing mechanism for online continual learning in transformers without forgetting, addressing non-stationary streaming data.

Ax Maximilian Beck, Jonas Gehring, Jannik Kossen, Gabriel Synnaeve 3/11/2026

Towards a Neural Debugger for Python

Neural debugger for Python that trains LLMs on execution traces to predict line-by-line program execution for debugging workflows.

Ax Ayushi Agarwal 3/11/2026

On the Formal Limits of Alignment Verification

Formal analysis proving that no verification procedure can simultaneously satisfy soundness, completeness, and decidability for AI alignment certification.

Ax Cornelius Emde, Alexander Rubinstein, Anmol Goel, Ahmed Heakl, Sangdoo Yun, Seong Joon Oh, Martin Gubri 3/11/2026

MASEval: Extending Multi-Agent Evaluation from Models to Systems

MASEval extends multi-agent evaluation beyond model-centric benchmarks to evaluate LLM-based agentic system components including topology, orchestration, and error handling.