Ax Jinwoo Kim, S\'ekou-Oumar Kaba, Jiyun Park, Seunghoon Hong, Siamak Ravanbakhsh 6/1/2026

Inverting Data Transformations via Diffusion Sampling

Probabilistic approach to inverting unknown group transformations using diffusion sampling for transformation recovery on Lie groups.

Ax Iv\'an Arcuschin, David Chanin, Adri\`a Garriga-Alonso, Oana-Maria Camburu 6/1/2026

Biases in the Blind Spot: Detecting What LLMs Fail to Mention

Proposes automated pipeline for detecting unverbalized biases in LLM chain-of-thought reasoning without requiring predefined categories or hand-crafted datasets.

Ax Tessa Han, Sebastian Bordt, Hanlin Zhang, Sham Kakade 6/1/2026

Weight Decay Improves Language Model Plasticity

Studies how weight decay during pretraining improves LLM plasticity and downstream adaptability beyond validation loss optimization.

Ax Yuxiang Guo, Zhuoran Du, Nan Tang, Kezheng Tang, Congcong Ge, Yunjun Gao 6/1/2026

DTBench: A Synthetic Benchmark for Document-to-Table Extraction

DTBench introduces a synthetic benchmark for evaluating LLM performance on document-to-table extraction tasks requiring complex reasoning and structured output generation.

Ax Kiho Park, Todd Nief, Yo Joong Choe, Victor Veitch 6/1/2026

The Information Geometry of Softmax: Probing and Steering

Studies information geometry of softmax representations in AI systems, focusing on how models encode semantic structure into representation spaces for behavior production.

Ax Yufei Li, Yisen Gao, Jiaxuan Xiong, Jiaxin Bai, Shijie Zhong, Haoyu Huang, Zhongwei Xie, Hong Ting Tsang, Yangqiu Song 6/1/2026

NGDBench: Towards Neural Graph Data Management

NGDBench proposes next-generation neural data management systems supporting heterogeneous evolving data with implicit reasoning.

Ax Charles Ye, Jasmine Cui, Dylan Hadfield-Menell 6/1/2026

Prompt Injection as Role Confusion

Research on prompt injection attacks as role confusion, showing LLMs identify text source by style rather than role labels, with measurement techniques.

Ax Rajinder Sandhu, Di Mu, Cheng Chang, Md Shahriar Tasjid, Himanshu Rai, Maksims Volkovs, Ga Wu 6/1/2026

Aligning Dense Retrievers with LLM Utility via Distillation

Utility-Aligned Embeddings framework improving dense retrieval for RAG by distilling LLM utility signals without expensive re-ranking.

Ax Bowen Zheng, Weijian Luo, Guang Yang, Colin Zhang, Tianyang Hu 6/1/2026

Autoregressive Visual Generation Needs a Prologue

Prologue approach bridging reconstruction-generation gap in autoregressive image generation by using separate tokens for each objective.

Ax Asher Labovich, Benjamin Bradley, Vanessa Alexander, Chaitanya Harsha 6/1/2026

Block-Based Double Decoders

Proposes block-based double decoders architecture combining decoder-only and encoder-decoder benefits with full loss supervision and efficient sequence packing.

Ax Max Prior, Natalia Milanova, Andreas Schultz 6/1/2026

Chunking German Legal Code

Compares chunking strategies for RAG on German legal code, benchmarking structural, semantic, and hierarchical retrieval approaches.

Ax Kritee Kondapally, Claire J. Smerdon, Pooja C. Patel, Ogheneyoma Akoni, Jevon Torres, Jaspreet Ranjit, Matthew Finlayson, Swabha Swayamdipta 6/1/2026

Side-by-side Comparison Amplifies Dialect Bias in Language Models

Studies covert dialect bias in language models, showing side-by-side comparisons amplify disparities in how LMs associate traits with dialectal variations.