Ax Maohao Shen, Tejas Jayashankar, Osama Hanna, Naoyuki Kanda, Yancheng Wang, Kate\v{r}ina \v{Z}mol\'ikov\'a, Ruiming Xie, Niko Moritz, Anfeng Xu, Yashesh Gaur, Gregory Wornell, Qing He, Jilong Wu 2/17/2026

GSRM: Generative Speech Reward Model for Speech RLHF

Generative Speech Reward Model for evaluating and improving naturalness in speech language model outputs via RLHF.

Ax Taiwei Shi, Sihao Chen, Bowen Jiang, Linxin Song, Longqi Yang, Jieyu Zhao 2/17/2026

Experiential Reinforcement Learning

Experiential RL training paradigm embedding explanatory chains for LMs to learn from sparse delayed environmental feedback.

Ax Dan Zhang, Yishu Lei, Jing Hu, Shuwei He, Songhe Deng, Xianlong Luo, Danxiang Zhu, Shikun Feng, Rui Liu, Jingzhou He, Yu Sun, Hua Wu, Haifeng Wang 2/17/2026

Eureka-Audio: Triggering Audio Intelligence in Compact Language Models

Eureka-Audio: 1.7B parameter compact audio language model matching performance of 7B-30B models on ASR and audio understanding.

Ax Yi Li, Hongze Shen, Lexiang Tang, Xin Li, Xinpeng Ding, Yinsong Liu, Deqiang Jiang, Xing Sun, Xiaomeng Li 2/17/2026

DenseMLLM: Standard Multimodal LLMs are Intrinsic Dense Predictors

DenseMLLM enables multimodal LLMs to perform dense prediction tasks like semantic segmentation and depth estimation without task-specific decoders.

Ax Nima Esmi (Bernoulli Institute, RUG, Groningen, Netherlands, ISRC, Khazar University, Baku, Azerbaijan), Maryam Nezhad-Moghaddam (Department of Computer Engineering, University of Guilan, Rasht, Iran), Fatemeh Borhani (Department of Computer Engineering, University of Guilan, Rasht, Iran), Asadollah Shahbahrami (ISRC, Khazar University, Baku, Azerbaijan, Department of Computer Engineering, University of Guilan, Rasht, Iran), Amin Daemdoost (Department of Computer Engineering, University of Guilan, Rasht, Iran), Georgi Gaydadjiev (QCE Department, TU Delft, Delft, Netherlands) 2/17/2026

GPT-5 vs Other LLMs in Long Short-Context Performance

Evaluates context window utilization across LLMs including GPT-5, comparing theoretical capacity vs practical performance on long-context tasks requiring detailed understanding.

Ax Yaxuan Kong, Hoyoung Lee, Yoontae Hwang, Alejandro Lopez-Lira, Bradford Levy, Dhagash Mehta, Qingsong Wen, Chanyeol Choi, Yongjae Lee, Stefan Zohren 2/17/2026

Evaluating LLMs in Finance Requires Explicit Bias Consideration

Analysis identifying five recurring biases in financial LLM evaluations: look-ahead, survivorship, narrative, objective, and cost.

Ax Tingting Tang, James Flemings, Yongqin Wang, Murali Annavaram 2/17/2026

Differentially Private Retrieval-Augmented Generation

Methods for differentially private retrieval-augmented generation protecting sensitive data while reducing hallucinations in LLM outputs.