Ax Xiaojie Xu, Zongyuan Li, Chang Lu, Runnan Qi, Yanan Ni, Lumin Jiang, Xiangbei Liu, Xuebo Zhang, Yongchun Fang, Kuihua Huang, Xian Guo, Zhanghua Wu, Zhenya Li 4/13/2026

Reflection of Episodes: Learning to Play Game from Expert and Self Experiences

Framework enabling LLMs to learn complex game strategies through self-reflection on expert and self-generated experiences in StarCraft II.

Ax Shahab Rahimirad, Guven Gergerli, Lucia Romero, Angela Qian, Matthew Lyle Olson, Simon Stepputtis, Joseph Campbell 4/13/2026

Bayesian Social Deduction with Graph-Informed Language Models

Study evaluating LLM performance on social reasoning tasks in the Avalon game, testing inference capabilities and model distillation effects.

Ax Zhenfeng Lin, Haoji Hu, Ming Hao, Xuchao Zhang, Ryan Zhang, Junhao Li, Ze Li, Oleg Kulygin, Chetan Bansal, Hatay Tuna, Murali Chintalapati, Sheila Jiang, Salman Zafar, Angie Anderson 4/13/2026

ActionNex: A Virtual Outage Manager for Cloud Computing

Production agentic system for cloud outage management with real-time updates, knowledge distillation, and conditioned action recommendations.

Ax Jingyang Qiao, Weicheng Meng, Yu Cheng, Zhihang Lin, Zhizhong Zhang, Xin Tan, Jingyu Gong, Kun Shao, Yuan Xie 4/13/2026

Memory Intelligence Agent

Memory system for deep research agents enabling efficient evolution and reasoning through intelligent trajectory memory management.

Ax Wenxuan Liu, Zixuan Li, Long Bai, Chunmao Zhang, Fenghui Zhang, Zhuo Chen, Wei Li, Yuxin Zuo, Fei Wang, Bingbing Xu, Xuhui Jiang, Jin Zhang, Xiaolong Jin, Jiafeng Guo, Tat-Seng Chua, Xueqi Cheng 4/13/2026

Towards Knowledgeable Deep Research: Framework and Benchmark

Framework and benchmark for deep research agents using structured knowledge alongside unstructured web content for comprehensive reports.

Ax Monishwaran Maheswaran, Leon Lakhani, Zhongzhu Zhou, Shijia Yang, Junxiong Wang, Coleman Hooper, Yuezhou Hu, Rishabh Tiwari, Jue Wang, Harman Singh, Qingyang Wu, Yuqing Jian, Ce Zhang, Kurt Keutzer, Tri Dao, Xiaoxia Wu, Ben Athiwaratkun, James Zou, Chenfeng Xu 4/13/2026

Squeeze Evolve: Unified Multi-Model Orchestration for Verifier-Free Evolution

Multi-model orchestration framework for verifier-free evolutionary inference balancing diversity and computational efficiency.

Ax Zixuan Hu, Yongxian Wei, Li Shen, Zhenyi Wang, Baoyuan Wu, Chun Yuan, Dacheng Tao 4/13/2026

Task-Distributionally Robust Data-Free Meta-Learning

Data-free meta-learning from pre-trained models without original training data, analyzing robustness and failure modes.

Ax Sajib Kumar Saha Joy, Arman Hassan Mahy, Meherin Sultana, Azizah Mamun Abha, MD Piyal Ahmmed, Yue Dong, G M Shahariar 4/13/2026

Mitigating Extrinsic Gender Bias for Bangla Classification Tasks

Investigation of gender bias in Bangla language models with benchmark datasets for sentiment analysis, toxicity detection, hate speech, and sarcasm.

Ax Jing-En Huang, I-Sheng Fang, Tzuhsuan Huang, Yu-Lun Liu, Chih-Yu Wang, Jun-Cheng Chen 4/13/2026

Gen-n-Val: Agentic Image Data Generation and Validation

Agentic framework for synthetic image data generation and validation addressing data scarcity and label noise in vision tasks like detection and segmentation.

Ax Alexander Gambashidze, Li Pengyi, Matvey Skripkin, Andrey Galichin, Anton Gusarov, Konstantin Sobolev, Andrey Kuznetsov, Ivan Oseledets 4/13/2026

Listener-Rewarded Thinking in VLMs for Image Preferences

Listener-rewarded thinking approach using reinforcement learning to train robust reward models for generative text-to-image and video models.