Ax Xavier Theimer-Lienhard, Mushtaha El-Amin, Fay Elhassan, Sahaj Vaidya, Victor Cartier-Negadi, David Sasu, Lars Klein, Mary-Anne Hartley 5/18/2026

Fully Open Meditron: An Auditable Pipeline for Clinical LLMs

Fully Open Meditron: Clinical LLM with complete transparent training pipeline including data provenance and curation procedures.

Ax Abinand Nallathambi, Christopher Knight, Shantanu Ganguly, Wilfried Haensch, Anand Raghunathan 5/18/2026

A3D: Agentic AI flow for autonomous Accelerator Design

A3D: Agentic AI system for autonomous hardware accelerator design using LLMs and tool integration to reduce design labor.

Ax Yihong Dong, Jiaru Qian, Haoran Zhang, Peixu Wang, Binhua Li, Zhi Jin, Yongbin Li, Ge Li, Xiaokang Yang, Xue Jiang 5/18/2026

From I/O to Code with Discovery Agent

Discovery agent system for automatic program synthesis from input-output behavior using LLMs.

Ax Sidharth Pulipaka, Stanislau Hlebik, Leonidas Raghav, Sahar Abdelnabi, Vyas Raina, Ivaxi Sheth, Mario Fritz 5/18/2026

Hidden in Memory: Sleeper Memory Poisoning in LLM Agents

Security analysis of sleeper memory poisoning attacks on LLM agents with persistent external memory.

Ax Jinyuan Li, Langlin Huang, Chengsong Huang, Shaoyang Xu, Donghong Cai, Yuyi Yang, Wenxuan Zhang, Jiaxin Huang 5/18/2026

Process Rewards with Learned Reliability

BetaPRM: Process Reward Model predicting step-level success probability and reliability scores for LLM reasoning verification.

Ax Haizhong Zheng, Yizhuo Di, Jiahui Wang, Shuowei Jin, Xueshen Liu, Yongji Wu, Z. Morley Mao, Ion Stoica, Jiawei Zhao, Beidi Chen 5/18/2026

AstraFlow: Dataflow-Oriented Reinforcement Learning for Agentic LLMs

AstraFlow: dataflow-oriented reinforcement learning system for scaling agentic LLMs with multi-policy training on heterogeneous, elastic compute resources.

Ax Shaoke Xi, ChonLam Lao, Boyi Jia, Jiaqi Gao, Zhipeng Zhang, Jiamin Cao, Brian Sutioso, Erci Xu, Minlan Yu, Kui Ren, Yong Li, Zhengping Qian, Ennan Zhai, Jingren Zhou 5/18/2026

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

PrismLLM: tool for emulating large-scale LLM training on small GPU clusters to enable debugging and optimization without accessing production infrastructure.

Ax Ali J Alrasheed, Aryan Yazdan Parast, Basim Azam, James Bailey, Naveed Akhtar 5/18/2026

Latent Video Prediction Learns Better World Models

Systematic evaluation of latent video prediction models as world models, analyzing robustness of V-JEPA, VideoPrism, and VideoMAEv2 across five axes.

Ax Jaeseung Heo, Kyeongheung Yun, Youngbin Choi, Sehyun Hwang, Jungseul Ok, Dongwoo Kim 5/18/2026

Interaction-Aware Influence Functions for Group Attribution

Interaction-aware influence functions measuring how groups of training examples jointly affect model predictions, capturing redundancy and complementarity.