HN meysamazad 5/18/2026

LLM Tracing with MLflow AI Gateway

MLflow AI Gateway tutorial on tracing agentic app interactions, token consumption, and tool usage for LLM debugging.

HN eigenBasis 5/18/2026

How to Work and Compound with AI

Workflow framework for iterative work with AI, focusing on context compounding and error reduction across sessions.

Ax Ming Yang, Zhiwei Zhang, Jiahang Li, Haoseng Liu, Yuzheng Cai, Weiguo Zheng 5/18/2026

DeepSlide: From Artifacts to Presentation Delivery

DeepSlide multi-agent system generates presentations with focus on delivery process including narrative planning and presentation preparation, not just artifacts.

Ax Jianbo Lin, Xiaomin Yu, Yi Xin, Yifu Guo, Zhuosong Jiang, Zhongqi Yue, Weishi Wang, Heqing Zou, Chengwei Qin, Hui Xiong 5/18/2026

ICRL: Learning to Internalize Self-Critique with Reinforcement Learning

ICRL uses reinforcement learning to help LLM agents internalize critique feedback, enabling improvement that persists without explicit critic guidance.

Ax Jingjing Wang, Xiwen Chen, Wenhui Zhu, Huayu Li, Zhengxiao He, Feiyang Cai, Ana S. Carreon-Rascon, Xuanzhao Dong, Feng Luo 5/18/2026

Context Pruning for Coding Agents via Multi-Rubric Latent Reasoning

Context pruning technique for coding agents using multi-rubric latent reasoning to reduce irrelevant repository files in context, improving token efficiency.

Ax Kin Max Piamolini Gusm\~ao, Nathan Gavenski, Nir Oren, Felipe Meneguzzi 5/18/2026

Zero-Shot Goal Recognition with Large Language Models

First systematic study of zero-shot goal recognition using LLMs, showing LLMs better suited for abductive consistency evaluation than novel plan generation.

Ax Fangzhou Lin, Shuo Xing, Peiran Li, Siyuan Yang, Qianwen Ge, Kazunori Yamada, Ziming Zhang, Haichong Zhang, Zhengzhong Tu 5/18/2026

CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning

CAPS optimizes parallel reasoning in LLMs by adaptively selecting which pairwise verification judgments to perform, reducing computational cost while maintaining solution quality.

Ax Alberto Pepe, Chien-Yu Lin, Despoina Magka, Bilge Acun, Yannan Nellie Wu, Anton Protopopov, Carole-Jean Wu, Yoram Bachrach 5/18/2026

Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design

AIRA framework with dual LLM agents autonomously designing neural architectures beyond Transformers through architecture search and mechanistic implementation.

Ax Michael Solodko, Justin Wagle 5/18/2026

ScreenSearch: Uncertainty-Aware OS Exploration

ScreenSearch system for GUI agents that explores desktop OS state under partial observability, reducing ambiguity before committing to actions.