Ax Marcel Hussing, Liv G. d'Aliberti, Claas Voelcker, Benjamin Eysenbach, Eric Eaton 5/22/2026

Behavior-Consistent Deep Reinforcement Learning

Addresses policy divergence in RL by formalizing behavior-consistent training for reliable deployment.

Ax Jon Saad-Falcon, Avanika Narayan, Hakki Orhun Akengin, J. Wes Griffin, Herumb Shandilya, Adrian Gamarra Lafuente, Medhya Goel, Rebecca Joseph, Shlok Natarajan, Etash Kumar Guha, Shang Zhu, Ben Athiwaratkun, John Hennessy, Azalia Mirhoseini, Christopher R\'e 5/22/2026

Intelligence per Watt: Measuring Intelligence Efficiency of Local AI

Study measuring efficiency of small language models on local hardware, comparing energy consumption and performance to cloud-based alternatives.

Ax Nuoya Xiong, Yuhang Zhou, Hanqing Zeng, Zhaorun Chen, Furong Huang, Shuchao Bi, Lizhu Zhang, Zhuokai Zhao 5/22/2026

Token-Level LLM Collaboration via FusionRoute

FusionRoute enables token-level collaboration between specialized LLMs to achieve broad domain performance without expensive scaling.

Ax Zhanming Shen, Jiaqi Hu, Zeyu Qin, Hao Chen, Wentao Ye, Zenan Huang, Yihong Zhuang, Guoshan Lu, Junlin Zhou, Junbo Zhao 5/22/2026

Training-Trajectory-Aware Token Selection

Training method for efficient LLM distillation addressing performance degradation in student models with strong reasoning ability.

Ax Elias J\"a\"asaari, Ville Hyv\"onen, Teemu Roos 5/22/2026

LEMUR: Learned Multi-Vector Retrieval

Proposes LEMUR, learned multi-vector retrieval system improving on ColBERT's late interaction model for information retrieval efficiency.

Ax Paul P. Hager, Fabian N. Harang, Luca Pelizzari, Samy Tindel 5/22/2026

The Volterra signature

Proposes Volterra signature as explicit feature representation for history-dependent systems, alternative to implicit memory mechanisms in RNNs and transformers.

Ax Tom Sander, Hongyan Chang, Tom\'a\v{s} Sou\v{c}ek, Tuan Tran, Valeriu Lacatusu, Sylvestre-Alvise Rebuffi, Alexandre Mourachko, Surya Parimi, Christophe Ropers, Rashel Moritz, Vanessa Stark, Hady Elsahar, Pierre Fernandez 5/22/2026

TextSeal: A Localized LLM Watermark for Provenance & Distillation Protection

TextSeal watermarking technique for LLMs based on Gumbel-max sampling with zero inference overhead and speculative decoding support.

Ax Fengfei Yu, Ruijia Niu, Dongxia Wu, Yian Ma, Rose Yu 5/22/2026

Calibrating LLMs with Semantic-level Reward

Method for calibrating LLMs with semantic-level rewards to improve uncertainty estimation in high-stakes applications.

Ax Akshay Manglik (Emily), Apaar Shanker (Emily), Kaustubh Deshpande (Emily), Jason Qin (Emily), Yash Maurya (Emily), Veronica Chatrath (Emily), Vijay S. Kalmath (Emily), Levi Lentz (Emily), Yuan (Emily), Xue 5/22/2026

Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents

Framework for systematic corpus-level diagnostics of LLM agent execution traces and failure patterns.

HN def- 5/22/2026

Finding Bugs Using LLMs

Materialize uses LLM-based coding agents with Claude to find bugs in code and pull requests, sharing implementation considerations and lessons learned.

HN mooreds 5/22/2026

The AI-Native Interview

Coding agents like Codex and Claude Code are transforming software engineering. Engineers must shift focus from writing precise code to designing agents that produce correct outcomes.

HN diebillionaires 5/22/2026

The AI Bubble – No One's Happy

OpenAI's CFO discusses financing chip and data center commitments through banks, private equity, and federal backing at WSJ Tech Live event.

HN c4pt0r 5/22/2026

Yet Another AI Teammate

Curated directory of AI teammate products including Perplexity Computer, a multi-model AI agent workspace for research, coding, data analysis, and workflow automation.