Ax Sayan Deb Sarkar, R\'emi Pautrat, Ondrej Miksik, Marc Pollefeys, Iro Armeni, Mahdi Rad, Mihai Dusmanu 2/16/2026

CoPE-VideoLM: Codec Primitives For Efficient Video Language Models

CoPE-VideoLM uses codec primitives for efficient video understanding in language models, reducing computational overhead while preserving temporal details.

Ax Weishun Zhong, Doron Sivan, Tankut Can, Mikhail Katkov, Misha Tsodyks 2/16/2026

Semantic Chunking and the Entropy of Natural Language

Statistical model capturing multi-scale structure of natural language entropy, benchmarking LLM compression rates against information-theoretic limits.

Ax Dheeraj Vattikonda, Santhoshi Ravichandran, Emiliano Penaloza, Hadi Nekoei, Megh Thakkar, Thibault Le Sellier de Chezelles, Nicolas Gontier, Miguel Mu\~noz-M\'armol, Sahar Omidi Shayegan, Stefania Raimondo, Xue Liu, Alexandre Drouin, Laurent Charlin, Alexandre Pich\'e, Alexandre Lacoste, Massimo Caccia 2/16/2026

How to Train Your LLM Web Agent: A Statistical Diagnosis

Statistical diagnosis of LLM web agents identifies bottlenecks in multi-step interactions and proposes methods reducing compute costs for open-source systems.

Ax Dhruv Jain, Harshit Shukla, Gautam Rajeev, Ashish Kulkarni, Chandra Khatri, Shubham Agarwal 2/16/2026

VoiceAgentBench: Are Voice Assistants ready for agentic tasks?

VoiceAgentBench benchmark evaluates speech language models on agentic tasks and adversarial robustness beyond isolated capabilities like transcription.

Ax Aboli Kathar, Aman Kumar, Anusha Kamath, Araveeti Srujan, Ashish Sharma, Chandra Bhushan, Divya Sorate, Duddu Prasanth Kumar, Evan Acharya, Harsh Sharma, Hrithik Kadam, Kanishk Singla, Keyur Doshi, Kiran Praveen, Kolisetty Krishna SK, Krishanu Adhikary, Lokesh MPT, Mayurdeep Sonowal, Nadeem Shaikh, Navya Prakash, Nimit Kothari, Nitin Kukreja, Prashant Devadiga, Rakesh Paul, Ratanjeet Pratap Chauhan, Raunak Kalani, Raviraj Joshi, Shamanth MH, Shantanu Pandey, Shubham Soni, Siddharth Dixit, Smriti Jopat, Sunil Patel, Suraj Singh, Suvradip Paul, Tulasi Pilla, Utkarsh Vaidya, Vineeth Nambiar, Vishal Kanvaty, Yatharth Dedhia 2/16/2026

FiMI: A Domain-Specific Language Model for Indian Finance Ecosystem

Domain-specialized financial language model for Indian digital payment systems adapted from Mistral architecture.

Ax Chengsong Huang, Wenhao Yu, Xiaoyang Wang, Hongming Zhang, Zongxia Li, Ruosen Li, Jiaxin Huang, Haitao Mi, Dong Yu 2/16/2026

R-Zero: Self-Evolving Reasoning LLM from Zero Data

Self-evolving LLM that autonomously generates and learns from reasoning tasks without human-curated data or labels.

Ax Milad Yazdani, Mahdi Mostajabdaveh, Zirui Zhou, Ying Xiong 2/16/2026

MASPRM: Multi-Agent System Process Reward Model

Multi-Agent System Process Reward Model that guides multi-agent inference through value assignment to partial transcripts using Monte Carlo Tree Search.

Ax Avi Bagchi, Akhil Bhimaraju, Moulik Choraria, Daniel Alabi, Lav R. Varshney 2/16/2026

Watermarking Discrete Diffusion Language Models

Watermarking technique for discrete diffusion language models to track AI-generated content and differentiate from human creations.

Ax Irina Saparina, Mirella Lapata 2/16/2026

Reasoning about Intent for Ambiguous Requests

Research on handling ambiguous requests in LLMs by generating multiple interpretation-answer pairs trained with RL and custom reward functions.