Ax Viacheslav Meshchaninov, Egor Shibaev, Artem Makoian, Ivan Klimov, Nikita Balagansky, Daniil Gavrilov, Aibek Alanov, Dmitry Vetrov 3/25/2026

Guided Star-Shaped Masked Diffusion

Novel sampling algorithm for masked diffusion models improving generation quality and efficiency.

Ax Bhavesh Kumar, Dylan Feng, Leonard Tang 3/25/2026

MJ1: Multimodal Judgment via Grounded Verification

arXiv paper MJ1: multimodal judge trained with RL enforcing visual grounding through structured verification chains and counterfactual consistency rewards.

Ax Sonia Laguna, Jorge da Silva Goncalves, Moritz Vandenhirtz, Alain Ryser, Irene Cannistraci, Julia E. Vogt 3/25/2026

Rethinking Machine Unlearning: Models Designed to Forget via Key Deletion

arXiv paper proposing key deletion approach for machine unlearning designed at model development stage rather than post-hoc, addressing privacy regulations and data errors.

Ax Chiyu Ma, Shuo Yang, Kexin Huang, Jinda Lu, Haoming Meng, Shangshang Wang, Bolin Ding, Soroush Vosoughi, Guoyin Wang, Jingren Zhou 3/25/2026

FIPO: Eliciting Deep Reasoning with Future-KL Influenced Policy Optimization

arXiv paper presenting FIPO reinforcement learning algorithm for improving reasoning in LLMs through fine-grained credit assignment beyond outcome-based rewards.

Ax Huancheng Chen, Jingtao Li, Weiming Zhuang, Chen Chen, Lingjuan Lyu 3/25/2026

Replay-Free Continual Low-Rank Adaptation with Dynamic Memory

Continual learning technique combining parameter-efficient fine-tuning with vision transformers to prevent catastrophic forgetting. Addresses sequential task adaptation.

Ax Riccardo Bravin, Massimo Pavan, Hazem Hesham Yousef Shalby, Fabrizio Pittorino, Manuel Roveri 3/25/2026

EmbBERT: Attention Under 2 MB Memory

Transformer attention mechanism compressed to run in under 2MB memory for IoT and wearable devices. Enables NLP deployment on ultra-constrained hardware.

Ax Raj Ghugare, Roger Creus Castanyer, Catherine Ji, Kathryn Wantlin, Jin Schofield, Karthik Narasimhan, Benjamin Eysenbach 3/25/2026

BuilderBench: The Building Blocks of Intelligent Agents

BuilderBench benchmark for evaluating intelligent agents' ability to learn through interaction and exploration beyond training data.

Ax Amos Goldman (NVIDIA Corporation), Nimrod Boker (NVIDIA Corporation), Maayan Sheraizin (NVIDIA Corporation), Nimrod Admoni (NVIDIA Corporation), Artem Polyakov (NVIDIA Corporation), Subhadeep Bhattacharya (NVIDIA Corporation), Fan Yu (NVIDIA Corporation), Kai Sun (NVIDIA Corporation), Georgios Theodorakis (NVIDIA Corporation), Hsin-Chun Yin (NVIDIA Corporation), Peter-Jan Gootzen (NVIDIA Corporation), Aamir Shafi (NVIDIA Corporation), Assaf Ravid (NVIDIA Corporation), Salvatore Di Girolamo (NVIDIA Corporation), Manjunath Gorentla Venkata (NVIDIA Corporation), Gil Bloch (NVIDIA Corporation) 3/25/2026

NCCL EP: Towards a Unified Expert Parallel Communication API for NCCL

NCCL EP presents a unified expert parallel communication API built on NCCL for GPU-initiated RDMA operations in Mixture-of-Experts LLM architectures.