Ax Jan Greb\'ik, Pavel Hub\'a\v{c}ek, Martin Kouteck\'y, Mat\v{e}j Kripner, V\'aclav Rozho\v{n}, Robert \v{S}\'amal, Adri\'an Z\'ame\v{c}n\'ik 4/21/2026

Bolzano: Case Studies in LLM-Assisted Mathematical Research

Bolzano multi-agent LLM system produces novel mathematics results through parallel provers and verifier with persistent knowledge base.

Ax Haocheng Ju, Leheng Chen, Peihao Wu, Bryan Dai, Bin Dong 4/21/2026

Matlas: A Semantic Search Engine for Mathematics

Semantic search engine for retrieving mathematical knowledge across millions of documents to ground AI mathematics systems.

Ax Feiyang Kang, Mahavir Dabas, Myeongseob Ko, Ruoxi Jia 4/21/2026

Characterizing Model-Native Skills

Research on characterizing language model skills through model-native representations rather than external taxonomies for intervention on model behavior.

Ax Prasoon Goyal, Sattvik Sahai, Michael Johnston, Hangjie Shi, Yao Lu, Shaohua Liu, Anna Rumshisky, Rahul Gupta, Anna Gottardi, Desheng Zhang, Lavina Vaz, Leslie Ball, Lucy Hu, Luke Dai, Samyuth Sagi, Maureen Murray, Sankaranarayanan Ananthakrishnan 4/21/2026

Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition

Dataset generation method using adversarial competition between attackers and defenders to create diverse, high-quality conversational training data.

Ax Wen Tao, Yiwei Wang, Peng Zhou, Bryan Hooi, Wanlong Fang, Tianle Zhang, Xiao Luo, Yuansheng Liu, Alvin Chan 4/21/2026

How Creative Are Large Language Models in Generating Molecules?

Research on using LLMs for molecular generation under chemical constraints, framing creativity as a functional requirement for exploring large chemical spaces.

Ax Michal Valko, R\'emi Munos, Branislav Kveton, Tom\'a\v{s} Koc\'ak 4/21/2026

Spectral bandits for smooth graph functions

Bandit algorithm framework for online learning on graph-structured payoffs with applications to content recommendation systems.

Ax Jiaqi Wang (Beijing University of Posts and Telecommunications, Beijing Academy of Artificial Intelligence), Haoge Deng (Beijing Academy of Artificial Intelligence), Ting Pan (Beijing Academy of Artificial Intelligence), Yang Liu (Beijing Academy of Artificial Intelligence), Chengyuan Wang (Beijing Academy of Artificial Intelligence), Fan Zhang (Beijing Academy of Artificial Intelligence), Yonggang Qi (Beijing University of Posts and Telecommunications), Xinlong Wang (Beijing Academy of Artificial Intelligence) 4/21/2026

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models

Framework for stable integration of reinforcement learning with uniform discrete diffusion models, addressing training instability in GRPO application.