Ax Huamin Chen, Xunzhuo Liu, Junchen Jiang, Bowei He, Xue Liu 4/17/2026

Token-Budget-Aware Pool Routing for Cost-Efficient LLM Inference

Token-budget-aware pool routing optimizes vLLM inference by routing requests based on estimated token requirements, reducing resource waste and KV-cache failures in production deployments.

Ax Hanchen David Wang, Diego Manzanas Lopez, Preston K. Robinette, Ipek Oguz, Taylor T. Johnson, Meiyi Ma 4/17/2026

Towards Verified and Targeted Explanations through Formal Methods

Formal methods combining XAI and neural network verification for trustworthy explanations with mathematical guarantees in safety-critical domains.

Ax Zhan Song, Yu-Tung Liu, Chen Chen, Guoheng Sun, Jiaqi Yin, Chia-tung Ho, Ang Li, Haoxing Ren, Cunxi Yu 4/17/2026

TOPCELL: Topology Optimization of Standard Cell via LLMs

TOPCELL framework using LLMs to optimize transistor topology in standard cell design, reformulating high-dimensional search as guided generation.

Ax Haoran Xu, Kaiwen Hu, Somayeh Sojoudi, Amy Zhang 4/17/2026

Reinforcement Learning via Value Gradient Flow

Behavior-regularized RL for LLM finetuning using value gradient flow to prevent value over-optimization from out-of-distribution extrapolation.

Ax Guillermo Valverde, Igor Garc\'ia-Olaizola, Giannicola Scarpa, Alejandro Pozas-Kerstjens 4/17/2026

Quantum-inspired tensor networks in machine learning models

Tensor network methods from quantum physics integrated into ML models to mitigate exponential complexity via compressed representations.

Ax Cheng Lu, Mengxin Wang, Dennis J. Zhang, Heng Zhang 4/17/2026

Generative Augmented Inference

Framework leveraging LLM outputs as auxiliary data for parameter estimation in operations management tasks.

Ax Minhak Song, Liang Zhang, Bingcong Li, Niao He, Michael Muehlebach, Sewoong Oh 4/17/2026

Zeroth-Order Optimization at the Edge of Stability

Zeroth-order optimization stability analysis for gradient-free learning and memory-efficient model fine-tuning.

Ax Xiaoyi Dong, Xi Sheryl Zhang, Jian Cheng 4/17/2026

Mean Flow Policy Optimization

MeanFlow: Few-step flow models replace diffusion for efficient policy representation in online RL.

Ax Amy Rouillard, Sitwala Mundiab, Linda Camarab, Michael Cameron Gramaniec, Ziyaad Dangorc, Ismail Kallad, Shabir A. Madhic, Kajal Morarc, Marlvin T. Ncubec, Haroon Saloojeee, Bruce A. Bassett 4/17/2026

Can LLMs Score Medical Diagnoses and Clinical Reasoning as well as Expert Panels?

Evaluates LLM juries against expert clinician panels for scoring medical diagnoses and clinical reasoning on real hospital cases.

Ax Dongsheng Wang, Jinsen Zhang, Dawei Su, Hui Huang 4/17/2026

Improving Sparse Autoencoder with Dynamic Attention

Proposes dynamic attention mechanism for sparse autoencoders to improve interpretability of foundation model activations by optimizing sparsity levels per neuron.