Ax Afonso Simpl\'icio, Gon\c{c}alo Vinagre, Miguel Moura Ramos, Diogo Tavares, Rafael Ferreira, Giuseppe Attanasio, Duarte M. Alves, In\^es Calvo, In\^es Vieira, Rui Guerra, James Furtado, Beatriz Canaverde, Iago Paulo, Vasco Ramos, Diogo Gl\'oria-Silva, Miguel Faria, Marcos Treviso, Daniel Gomes, Pedro Gomes, David Semedo, Andr\'e Martins, Jo\~ao Magalh\~aes 3/30/2026

AMALIA Technical Report: A Fully Open Source Large Language Model for European Portuguese

AMALIA: fully open-source LLM trained on high-quality European Portuguese data with native evaluation benchmark.

Ax James A. Michaelov, Catherine Arnett, Tyler A. Chang, Pamela D. Rivi\`ere, Samuel M. Taylor, Cameron R. Jones, Sean Trott, Roger P. Levy, Benjamin K. Bergen, Micah Altman 3/30/2026

How Open Must Language Models be to Enable Reliable Scientific Inference?

Analysis of how open vs closed language models impact scientific inference reliability and reproducibility.

Ax Shihua Zhang, Qiuhong Shen, Shizun Wang, Tianbo Pan, Xinchao Wang 3/30/2026

Make Geometry Matter for Spatial Reasoning

Study on improving vision-language models' spatial reasoning by injecting geometry tokens from 3D foundation models.

Ax Shaoxuan Li, Zhixuan Zhao, Hanze Deng, Zirun Ma, Shulin Tian, Zuyan Liu, Yushi Hu, Haoning Wu, Yuhao Dong, Benlin Liu, Ziwei Liu, Ranjay Krishna 3/30/2026

PerceptionComp: A Video Benchmark for Complex Perception-Centric Reasoning

PerceptionComp: benchmark for complex long-horizon perception-centric video reasoning requiring compositional temporal evidence and multi-subtask integration.

Ax Sijia Liu, Niklas Muennighoff, Kawin Ethayarajh 3/30/2026

Humanline: Online Alignment as Perceptual Loss

Online alignment (GRPO) outperforms offline alignment (DPO) through better approximation of human-perceived model output distribution based on prospect theory.

Ax Zhengru Fang, Yu Guo, Yuang Zhang, Haonan An, Wenbo Ding, Yuguang Fang 3/30/2026

Shared Spatial Memory Through Predictive Coding

Multi-agent predictive coding framework for constructing shared spatial memory through mutual uncertainty minimization with information bottleneck objective.

Ax Yibo Yang, Fei Lei, Yixuan Sun, Yantao Zeng, Chengguang Lv, Jiancao Hong, Jiaojiao Tian, Tianyu Qiu, Xin Wang, Yanbing Chen, Yanjie Li, Zheng Pan, Xiaochen Zhou, Guanzhou Chen, Haoran Lv, Yuning Xu, Yue Ou, Haodong Liu, Shiqi He, Anya Jia, Yulei Xin, Huan Wu, Liang Liu, Jiaye Ge, Jianxin Dong, Dahua Lin, Wenxiu Sun 3/30/2026

AIDABench: AI Data Analytics Benchmark

AIDABench: comprehensive benchmark for evaluating AI-driven document understanding and processing tools in end-to-end real-world scenarios.

Ax Julia Jose, Meghna Manoj Nair, Rachel Greenstadt 3/30/2026

Large-Scale Analysis of Persuasive Content on Moltbook

NLP study detecting political propaganda on Moltbook platform using LLM-based classifiers on 673k posts; analyzes prevalence and concentration patterns.

Ax Akio Kodaira, Tingbo Hou, Ji Hou, Markos Georgopoulos, Felix Juefei-Xu, Masayoshi Tomizuka, Yue Zhao 3/30/2026

StreamDiT: Real-Time Streaming Text-to-Video Generation

StreamDiT enables real-time streaming text-to-video generation using transformer-based diffusion models.

Ax Hongxiang Zhang, Yuan Tian, Tianyi Zhang 3/30/2026

Attention-Aligned Reasoning for Large Language Models

arXiv research on ATAR method improving LLM reasoning by aligning attention with reasoning structure. Prevents critical steps from being buried in extended reasoning chains.

Ax T\'eo Guichoux, Th\'eodor Lemerle, Shivam Mehta, Jonas Beskow, Gustav Eje Henter, Laure Soulier, Catherine Pelachaud, Nicolas Obin 3/30/2026

Gelina: Unified Speech and Gesture Synthesis via Interleaved Token Prediction

arXiv research on unified speech and gesture synthesis via interleaved token prediction. Jointly generates synchronized co-speech gestures and speech from text with discrete autoregressive model.

Ax Minsuk Ji, Sanghyeok Lee, Namhyuk Ahn 3/30/2026

Compositional Image Synthesis with Inference-Time Scaling

arXiv research on training-free compositional image synthesis combining object-centric approaches with self-refinement. Uses LLMs to improve layout faithfulness in text-to-image generation.

Ax Tiansheng Wen, Yifei Wang, Aosong Feng, Long Ma, Xinyang Liu, Yifan Wang, Lixuan Guo, Bo Chen, Stefanie Jegelka, Chenyu You 3/30/2026

Route Experts by Sequence, not by Token

arXiv research on Sequence-level TopK routing for Mixture-of-Experts LLMs. Minimal modification enabling adaptive expert routing based on token complexity without retraining.

Ax Rongbin Hu, Jeffrey Liu 3/30/2026

Binary Verification for Zero-Shot Vision

arXiv research on training-free binary verification workflow for zero-shot vision using VLMs. Converts open-ended queries to multiple-choice with deterministic resolution.

Ax Zhenchao Tang, Fang Wang, Haohuai He, Jiale Zhou, Tianxu Lv, Jun Zhu, Shouzhi Chen, Minghao Yang, Yu Wang, Jiayang Wu, Yidong Song, Yaokun Li, Jiehui Huang, Dawei Huang, Zhi Song, Jianhua Yao 3/30/2026

Aligning LLMs with Biomedical Knowledge using Balanced Fine-Tuning

arXiv research proposing Balanced Fine-Tuning method for aligning LLMs with biomedical knowledge. Combines SFT and RL using confidence-weighted token optimization for scientific understanding.

Ax Woongyeong Yeo, Kangsan Kim, Jaehong Yoon, Sung Ju Hwang 3/30/2026

WorldMM: Dynamic Multimodal Memory Agent for Long Video Reasoning

arXiv research on multimodal memory architecture for long-form video understanding. Addresses context capacity and visual detail retention in hours-long videos using dynamic memory mechanisms.