Ax Jaechang Kim, Yotaro Shimose, Zhao Wang, Kuang-Da Wang, Jungseul Ok, Shingo Takamatsu 7/1/2026

Visual Prompt Discovery via Semantic Exploration

Method for generating visual prompts via semantic exploration to improve LLM perception and visual reasoning capabilities.

Ax Juhan Park, Taerim Yoon, Seungmin Kim, Joong-Gil Kim, Wontae Ye, Jeongeun Park, Yoonbyung Chai, Geonwoo Cho, Geunwoo Cho, Dohyeong Kim, Kyungjae Lee, Yong-Jae Kim, Sungjoon Choi 7/1/2026

Learning Dexterous Grasping from Sparse Taxonomy Guidance

Research on dexterous robotic manipulation using sparse taxonomy guidance for grasp planning with reinforcement learning and multi-finger control.

Ax Siyu Chen, Miao Lu, Beining Wu, Heejune Sheen, Fengzhuo Zhang, Shuangning Li, Zhiyuan Li, Jose Blanchet, Tianhao Wang, Zhuoran Yang 7/1/2026

INFUSER: Influence-Guided Self-Evolution Improves Reasoning

arXiv paper presenting INFUSER, iterative co-training framework for self-improving LLM reasoning without heavy supervision.

Ax Xilong Wang, Xiaoxing Chen, Patrick Li, Dawn Song, Neil Gong 7/1/2026

Same-Origin Policy for Agentic Browsers

arXiv paper studying security of agentic browsers, examining same-origin policy effectiveness with AI agents.

Ax Soham Bhattacharjee, Dushyant Singh Chauhan, Salem Lahlou, Martin Takac, Nils Lukas 7/1/2026

Entropy-Gated Latent Recursion

arXiv paper proposing inference-time scaling method for LLM reasoning via deterministic layer recursion.

Ax Annika Marie Schoene, Cansu Canca, Gautham Vijay Kumar, Anson Antony 7/1/2026

One Year Later...The Harms Persist, But So Do We!

Evaluates safety guardrails of eight LLMs across 16 psychiatric conditions using adversarial attacks and introduces harm taxonomy framework.

Ax Kan Zhu, Mathew Jacob, Chenxi Ma, Yi Pan, Stephanie Wang, Arvind Krishnamurthy, Baris Kasikci 7/1/2026

TraceLab: Characterizing Coding Agent Workloads for LLM Serving

TraceLab provides a public dataset and analysis of real coding agent workloads across multiple LLM-based agents and models to improve serving efficiency.

Ax Woernle Frank, Fedosov Vladimir, Grinenko Artemiy 7/1/2026

Hierarchical Global Attention (HGA)

Hierarchical Global Attention drop-in replacement for dense causal attention enabling 64K-token context without retraining or parameter changes.

Ax Rajat Ghosh, Datta Nimmaturi, Aryan Singhal, Vaishnavi Bhargava, Henry Wong, Johnu George, Debojyoti Dutta 7/1/2026

Predictable GRPO: A Closed-Form Model of Training Dynamics

First-principles reduced-order model of GRPO training dynamics for LLM reasoning with closed-form analysis replacing empirical hyperparameter tuning.

Ax Shreyas Rajesh, Kartik Sharma, Tonmoy Monsoor, Mehmet Yigit Turali, Richard Idro, Juliana Kayaga, Robert Sebunya, Tracy Tushabe Namata, Jessica Nichole Pasqua, Vwani Roychowdhury, Rajarshi Mazumder 7/1/2026

Teaching LLMs to Recommend and Defer in Underrepresented Epilepsy Care

LLM-based clinical decision support for epilepsy medication prediction in resource-constrained settings, adapted to local practice with deferral capability.

Ax Apurva Gandhi, Vishwas Suryanarayanan, Raja Hasnain Anwar, Firoz Shaik, Shubhang Desai, Thong Q. Nguyen, Muhammad Taqi Raza, Vishal Chowdhary, Graham Neubig 7/1/2026

PPT-Eval: A Benchmark for Computer-Use Agents on PowerPoint Tasks

Benchmark with 120 PowerPoint tasks evaluating computer-use AI agents on multimodal content creation and presentation editing scenarios.