Ax Mehran Taghian, Yunke Peng, Xing Huang, Yao Wang, Yaoyuan Wang, Wei Guo, Yuanyong Luo, Tianchi Hu, Junsong Wang, Xin Wang, Hu Liu, Yu Cheng, Ziwei Yu, Hongliang Li, Mehdi Rahimifar, Lei Yan, Xuefei Wang, Zhuang Ma, Lei Liu, Hui Yu, Anandharaju Durai Raju, Hoang Le, Hei Yi Mak, Tanzila Rahman, Shadan Golestan 4/13/2026

HiFloat4 Format for Language Model Pre-training on Ascend NPUs

4-bit floating-point format (HiFloat4) for efficient language model pre-training on Ascend NPU hardware.

Ax Chia-Hong Hsu, Frank Wood 4/13/2026

Discrete Meanflow Training Curriculum

Training curriculum method for discrete flow-based image generation models to improve one-step sampling stability and quality.

Ax Amrut Nadgir, Vijay Balasubramanian, Pratik Chaudhari 4/13/2026

How does Chain of Thought decompose complex tasks?

Demonstrates power-law scaling of classification error with number of classes and how chain-of-thought decomposition reduces error through task splitting.

Ax Maksim Anisimov (Imperial College London), Francesco Belardinelli (Imperial College London), Matthew Wicker (Imperial College London) 4/13/2026

SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning

Method for safely updating deep reinforcement learning policies while preserving safety guarantees on previously encountered tasks.

Ax Wenjie Qu, Xuandong Zhao, Jiaheng Zhang, Dawn Song 4/13/2026

Self-Sovereign Agent

Investigation of self-sovereign AI agents that can economically sustain themselves without human involvement using LLMs and agent frameworks.