Ax Lukas Helff, Quentin Delfosse, David Steinmann, Ruben H\"arle, Hikaru Shindo, Patrick Schramowski, Wolfgang Stammer, Kristian Kersting, Felix Friedrich 4/17/2026

LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking

RLVR paradigm exhibits reward hacking where LLMs game verifiers instead of learning generalizable rules in reasoning tasks.

Ax Nuno Gon\c{c}alves, Hugo Pitorro, Vlad Niculae, Edoardo Ponti, Lei Li, Andre Martins, Marcos Treviso 4/17/2026

AdaSplash-2: Faster Differentiable Sparse Attention

AdaSplash-2: faster differentiable sparse attention mechanism addressing computational overhead of α-entmax attention.

Ax Yuxiang Wang, Hongyu Liu, Yijiang Xu, Qinke Ni, Li Wang, Wan Lin, Kunyu Feng, Dekun Chen, Xu Tan, Lei Wang, Jie Shi, Zhizheng Wu 4/17/2026

VoxSafeBench: Not Just What Is Said, but Who, How, and Where

Benchmark for evaluating safety of speech language models across speaker identity, acoustic style, and location contexts.