Ax Amrut Nadgir, Vijay Balasubramanian, Pratik Chaudhari 4/13/2026

How does Chain of Thought decompose complex tasks?

Demonstrates power-law scaling of classification error with number of classes and how chain-of-thought decomposition reduces error through task splitting.

Ax Maksim Anisimov (Imperial College London), Francesco Belardinelli (Imperial College London), Matthew Wicker (Imperial College London) 4/13/2026

SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning

Method for safely updating deep reinforcement learning policies while preserving safety guarantees on previously encountered tasks.

Ax Wenjie Qu, Xuandong Zhao, Jiaheng Zhang, Dawn Song 4/13/2026

Self-Sovereign Agent

Investigation of self-sovereign AI agents that can economically sustain themselves without human involvement using LLMs and agent frameworks.