Ax Baoqing Yue, Zihan Zhu, Yifan Zhang, Jichen Feng, Hufei Yang, Mengdi Wang 3/6/2026

Interactive Benchmarks

Interactive Benchmarks: evaluation framework assessing model reasoning ability through active information acquisition under constraints.

Ax Yuchen Shi, Huajie Chen, Heng Xu, Zhiquan Liu, Jialiang Shen, Chi Liu, Shuai Zhou, Tianqing Zhu, Wanlei Zhou 3/6/2026

Osmosis Distillation: Model Hijacking with the Fewest Samples

Research on security threats in transfer learning combined with dataset distillation, proposing osmosis distillation for model hijacking with minimal samples.

Ax Kenan Li, Rongzhi Li, Linghao Zhang, Qirui Jin, Liao Zhu, Xiaosong Huang, Geng Zhang, Yikai Zhang, Shilin He, Chengxing Xie, Xin Zhang, Zijian Jin, Bowen Li, Chaoyun Zhang, Yu Kang, Yufan Huang, Elsie Nallipogu, Saravan Rajmohan, Qingwei Lin, Dongmei Zhang 3/6/2026

RepoLaunch: Automating Build&Test Pipeline of Code Repositories on ANY Language and ANY Platform

RepoLaunch agent automates dependency resolution, compilation, and test extraction for software repositories across languages and platforms.

Ax Anatoly Belikov, Ilya Fedotov 3/6/2026

Good-Enough LLM Obfuscation (GELO)

GELO obfuscation scheme protects LLM prompt privacy on shared accelerators against memory access attacks.

Ax Jonathan D. Chang, Andrew Drozdov, Shubham Toshniwal, Owen Oertell, Alexander Trott, Jacob Portes, Abhay Gupta, Pallavi Koppol, Ashutosh Baheti, Sean Kulinski, Ivan Zhou, Irene Dea, Krista Opsahl-Ong, Simon Favreau-Lessard, Sean Owen, Jose Javier Gonzalez Ortiz, Arnav Singhvi, Xabi Andrade, Cindy Wang, Kartik Sreenivasan, Sam Havens, Jialu Liu, Peyton DeNiro, Wen Sun, Michael Bendersky, Jonathan Frankle 3/6/2026

KARL: Knowledge Agents via Reinforcement Learning

KARL system trains enterprise search agents via reinforcement learning with KARLBench evaluation suite covering six search tasks including entity search and report synthesis.

Ax Robin Shing Moon Chan, Tianyu Liu, Samuel Kiegeland, Clemente Pasti, Jacob Hoover Vigly, Timothy J. O'Donnell, Ryan Cotterell, Tim Vieira 3/6/2026

Ensembling Language Models with Sequential Monte Carlo

arXiv paper on ensemble methods for language models using Sequential Monte Carlo to aggregate predictions from multiple LLMs and prompting strategies.

Ax Siddharth Boppana, Annabel Ma, Max Loeffler, Raphael Sarfati, Eric Bigelow, Atticus Geiger, Owen Lewis, Jack Merullo 3/6/2026

Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought

arXiv paper analyzing chain-of-thought reasoning in models, showing performative generation where models generate tokens without revealing internal beliefs.

Ax Fred Zhangzhi Peng, Zachary Bezemek, Sawan Patel, Jarrid Rector-Brooks, Sherwood Yao, Avishek Joey Bose, Alexander Tong, Pranam Chatterjee 3/6/2026

Path Planning for Masked Diffusion Model Sampling

Method for improving masked diffusion models with path planning to enable iterative refinement during token generation.

Ax Halil Alperen Gozeten, M. Emrullah Ildiz, Xuechen Zhang, Hrayr Harutyunyan, Ankit Singh Rawat, Samet Oymak 3/6/2026

Continuous Chain of Thought Enables Parallel Exploration and Reasoning

CoT2 extends chain-of-thought reasoning in LLMs using continuous token representations instead of discrete sampling, with theoretical guarantees for logical reasoning tasks.

Ax C\'edric L\'eonard (Technical University of Munich, Munich, Germany, Remote Sensing Technology Institute), Dirk Stober (Technical University of Munich, Munich, Germany), Martin Schulz (Technical University of Munich, Munich, Germany) 3/6/2026

FPGA-Enabled Machine Learning Applications in Earth Observation: A Systematic Review

Systematic review of FPGA-based ML deployment for real-time Earth observation processing on UAVs and satellite systems with bandwidth constraints.

Ax Keyue Jiang, Jiahao Cui, Xiaowen Dong, Laura Toni 3/6/2026

Bures-Wasserstein Flow Matching for Graph Generation

Develops Bures-Wasserstein flow matching for graph generation using optimal transport-based probability paths for drug discovery and circuit design.

Ax Luca Serfilippi, Giorgio Franceschelli, Antonio Corradi, Mirco Musolesi 3/6/2026

Complexity-Regularized Proximal Policy Optimization

Proposes complexity-regularized proximal policy optimization replacing entropy regularization with self-regulating complexity terms for policy gradient methods.