Ax Shi Feng, Hanlin Zhang, Fan Nie, Sham Kakade, Yiling Chen 19d ago

Peer-Predictive Self-Training for Language Model Reasoning

Peer-Predictive Self-Training (PST) framework enables collaborative self-improvement of language models without external supervision using cross-model aggregation.

Ax Bingxi Zhao, Jiahao Zhang, Xubin Ren, Zirui Guo, Tianzhe Chu, Yi Ma, Chao Huang 19d ago

DeepTutor: Towards Agentic Personalized Tutoring

Open-source agentic tutoring framework combining LLMs with personalized feedback, difficulty calibration, and citation-grounded problem solving.

Ax Liang Luo, Yinbin Ma, Quanyu Zhu, Vasiliy Kuznetsov, Yuxin Chen, Neng Shi, Jian Jiao, Jiecao Yu, Buyun Zhang, Tongyi Tang, Xiaohan Wei, Yanli Zhao, Zeliang Chen, Yuchen Hao, Venkatesh Ranganathan, Sandeep Parab, Yantao Yao, Maxim Naumov, Chunzhi Yang, Shen Li, Ellie Wen, Wenlin Chen, Santanu Kolay, Chunqiang Tang 19d ago

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale

Low-precision FP8 arithmetic optimization techniques for large recommendation models, addressing numerical sensitivity and training efficiency.

Ax Ian Rios-Sialer, Shantanu Darveshi, Shuai Jiang, Avigya Paudel, Anastasiia Pronina, Ipshita Bandyopadhyay, Justin Shenk 19d ago

Temporal Preference Concepts and their Functions in a Large Language Model

arXiv paper identifying and analyzing temporal preference subgraphs in LLMs through causal localization, showing how models represent temporal tradeoffs internally.

Ax Abhinav Agarwal, Adam Wei, Taylan Kargin, Michael Zeng, Cole Becker, Arif Kerem Dayi, Pablo Parrilo, Asuman Ozdaglar, Russ Tedrake 19d ago

Training and Evaluating Diffusion Policies with Long Context Lengths

Benchmark of diffusion policies for robotic manipulation with incrementally increasing context lengths to enable memory and long-horizon task performance.

Ax Jannik H\"osch, Alessandro Sestini, Florian Fuchs, Amir Baghi, Joakim Bergdahl, Iolanda Leite, Konrad Tollmar, Jean-Philippe Barrette-LaPierre, Linus Gissl\'en 19d ago

Hierarchical Control in Multi-Agent Games: LLM-based Planning and RL Execution

Hierarchical multi-agent architecture combining LLM-based strategic planning with specialized RL skill policies for complex coordinated decision-making.

Ax Yanis Labrak, Dairazalia Sanchez-Cortes, Sergio Burdisso, S\'everin Baroudi, Shashi Kumar, Esa\'u Villatoro-Tello, Srikanth Madikeri, Manjunath K E, Old\v{r}ich Plchot, Kadri Hacio\u{g}lu, Petr Motlicek, Andreas Stolcke 19d ago

How to Leverage Synthetic Speech for LLM-Based ASR Systems?

Study on leveraging synthetic TTS speech for training ASR systems in privacy-constrained domains like banking and healthcare, addressing synthetic-real data gaps.

Ax R\'ois\'in Luo, Christian Gagn\'e, Jonas Ngnaw\'e, Ihsan Ullah, Karyn Morrissey 19d ago

A Stochastic--Geometric Theory of Scaling Laws in Grokking

Theoretical characterization of grokking phenomenon using stochastic-geometric analysis of solution space topology and delayed generalization in neural networks.

Ax Zhuoxuan Zhang (Yang), Kangqi Ni (Yang), Yuhang Chen (Yang), Mingfu Liang (Yang), Xiaohan Wei (Yang), Yunchen Pu (Yang), Fei Tian (Yang), Chonglin Sun (Yang), Frank Shyu (Yang), Adam (Yang), Song, Sandeep Pandey, Luke Simon, Tianlong Chen, Xi Liu 19d ago

Diffusion-GR2: Diffusion Generative Reasoning Re-ranker

Diffusion-GR2 uses block-diffusion language models for faster generative reasoning re-ranking with parallel decoding instead of sequential autoregressive inference.

Ax Vaishnavi Sinha, Pooja Guttal, Pranay Deep Reddy Katike, Vishal Sinha, Gerald Ndawula, Lira Yoon, Andrea Kleinsmith, Manas Gaur 19d ago

Where do LLMs Fall Short in CBT-Guided Affective Reasoning?

Study analyzing LLM failures in applying Cognitive Behavioral Therapy frameworks despite high theoretical knowledge, highlighting reasoning gaps in practical application.

Ax Alexis Kafantaris 19d ago

LLM for the development of FCM

Approach using local LLMs to extract quantitative data from text for developing fuzzy cognitive maps.

Ax Anne Harrington, Nayan Saxena, Michael Murphy, Anastasia Borovykh, Zeyu Yun, Sridhar Kamath, Ara Eindra Kyi, Trevor Darrell, Jitendra Malik, Yutong Bai 19d ago

When Does Continual Learning Require Learning

Framework for continual learning in LLMs distinguishing between domain adaptation and competence improvement as world conditions change.