Ax Yuanhe Zhang, Xinyue Wang, Zhican Chen, Weiliu Wang, Zilu Zhang, Zhengshuo Gong, Zhenhong Zhou, Kun Wang, Li Sun, Yang Liu, Sen Su 3/19/2026

Resource Consumption Threats in Large Language Models

Survey of resource consumption threats in LLMs, covering efficiency issues affecting service capacity, latency, and API costs.

Ax Zhengbo Zhang, Jinbo Su, Zhaowen Zhou, Changtao Miao, Yuhan Hong, Qimeng Wu, Yumeng Liu, Feier Wu, Yihe Tian, Yuhao Liang, Zitong Shan, Wanke Xia, Yi-Fan Zhang, Bo Zhang, Zhe Li, Shiming Xiang, Ying Yan 3/19/2026

VisBrowse-Bench: Benchmarking Visual-Native Search for Multimodal Browsing Agents

Benchmark for visual-native search in multimodal browsing agents, evaluating MLLM visual reasoning over web pages.

Ax G. Ciarfaglia, A. Rosanova, S. Cipolla, J. Bartoli, A. Di Domenico, C. Fioroni, A. Fontana, M. R. Scoleri, M. I. Mone, D. Franchi, M. C. Del Gaudio, F. Picariello, M. Gabusi, S. Bonura, V. Morreale, I. Bailo 3/19/2026

EngGPT2: Sovereign, Efficient and Open Intelligence

Italian open-source LLM with 16B parameters achieving competitive performance on benchmarks while requiring fraction of inference power.

Ax Omer Nacar, Deema Alquffari, Saleh Alsharideh, Adeem AlOtaibi, Abdulaziz Alabdulkarim, Leen Alhazmi, Nada Alomar, Wareef Alzubaidi, Nada Alsultan, Ahmed Alrabghi, Demah Alhoshan, Rana Alsayyari, Hamed Alruwaili, Albaraa Jaafar, Khaled Alusmani, Abdulaziz Alsohimy, Munirah Alsubaie, Shahd Aldukhayil, Arwa Alali, Yazeed BinShihah, Razan Alsulaymi, Nourah Alhumaid, Razan Abdulsalam, Reem Alamoudi, Mohammed Alkhalifa 3/19/2026

From Language to Action in Arabic: Reliable Structured Tool Calling via Data-Centric Fine-Tuning

Production framework for reliable Arabic function-calling models enabling agentic AI systems through data-centric fine-tuning.

Ax Benjamin Hudson, Laurent Charlin, Emma Frejinger 3/19/2026

Contextual Preference Distribution Learning

Proposes pipeline to learn context-dependent preference distributions for risk-averse decision-making via inverse optimization.

Ax Peng Xia, Jianwen Chen, Xinyu Yang, Haoqin Tu, Jiaqi Liu, Kaiwen Xiong, Siwei Han, Shi Qiu, Haonian Ji, Yuyin Zhou, Zeyu Zheng, Cihang Xie, Huaxiu Yao 3/19/2026

MetaClaw: Just Talk -- An Agent That Meta-Learns and Evolves in the Wild

MetaClaw enables LLM agents to continuously adapt and evolve in production by meta-learning from diverse task distributions without storing raw trajectories.

Ax Seyed Mohammad Asghari, Chris Chute, Vikranth Dwaracherla, Xiuyuan Lu, Mehdi Jafarnia, Victor Minden, Zheng Wen, Benjamin Van Roy 3/19/2026

Efficient Exploration at Scale

arXiv: Online learning algorithm for RLHF that improves data efficiency. Incrementally updates reward and language models from choice data.

Ax Dilxat Muhtar, Jiashun Liu, Wei Gao, Weixun Wang, Shaopan Xiong, Ju Huang, Siran Yang, Wenbo Su, Jiamang Wang, Ling Pan, Bo Zheng 3/19/2026

Complementary Reinforcement Learning

arXiv paper on complementary reinforcement learning for LLM-based agents, improving sample efficiency by leveraging historical experience across episodes.