Ax Nicholas Kuhn, Arvid Weyrauch, Lars Heyen, Achim Streit, Markus G\"otz, Charlotte Debus 2/24/2026

Bayesian Lottery Ticket Hypothesis

Applies lottery ticket hypothesis to Bayesian neural networks, finding sparse subnetworks for uncertainty quantification with reduced computational cost.

Ax Peter Romero, Fernando Mart\'inez-Plumed, Zachary R. Tyler, Matthieu T\'eh\'enan, Sipeng Chen, \'Alvaro David G\'omez Ant\'on, Luning Sun, Manuel Cebrian, Lexin Zhou, Yael Moros Daval, Daniel Romero-Alvarado, F\'elix Mart\'i P\'erez, Kevin Wei, Jos\'e Hern\'andez-Orallo 2/24/2026

From Human-Level AI Tales to AI Leveling Human Scales

Framework for calibrating AI benchmark performance against world population baselines to provide human-anchored capability scales.

Ax David Li, Nikita Gushchin, Dmitry Abulkhanov, Eric Moulines, Ivan Oseledets, Maxim Panov, Alexander Korotin 2/24/2026

IDLM: Inverse-distilled Diffusion Language Models

Inverse distillation for diffusion language models. Accelerates discrete diffusion models for faster text generation inference.

Ax Abhinav Moudgil, Boris Knyazev, Eugene Belilovsky 2/24/2026

Celo2: Towards Learned Optimization Free Lunch

Celo2 learned optimizer with improved meta-generalization. Aims for practical adoption of learned optimization rules beyond hand-designed optimizers.

Ax Teresa Yeo, Myeongho Jeon, Dulaj Weerakoon, Rui Qiao, Alok Prakash, Armando Solar-Lezama, Archan Misra 2/24/2026

Adaptive Problem Generation via Symbolic Representations

Adaptive problem generation via symbolic representations for training small open-weight LMs on math tasks. Data generation using RL with verifiable rewards.

Ax Egor Denisov, Svetlana Glazyrina, Maksim Kryzhanovskiy, Roman Ischenko 2/24/2026

Smooth Gate Functions for Soft Advantage Policy Optimization

Proposes Soft Adaptive Policy Optimization (SAPO) replacing hard clipping with smooth sigmoid gate functions to stabilize LLM training and reasoning in GRPO framework.

Ax Daniel Ritter, Owen Oertell, Bradley Guo, Jonathan Chang, Kiant\'e Brantley, Wen Sun 2/24/2026

LLMs Can Learn to Reason Via Off-Policy RL

Method for training LLMs to reason using off-policy reinforcement learning, addressing policy lag in distributed training architectures.

Ax Zelin He, Boran Han, Xiyuan Zhang, Shuai Zhang, Haotian Lin, Qi Zhu, Haoyang Fang, Danielle C. Maddix, Abdul Fatir Ansari, Akash Chandrayan, Abhinav Pradhan, Bernie Wang, Matthew Reimherr 2/24/2026

SenTSR-Bench: Thinking with Injected Knowledge for Time-Series Reasoning

Benchmark combining general reasoning LLMs with domain-specific time-series knowledge for improved time-series diagnostic reasoning tasks.

Ax Bryan Guanrong Shan, Alysa Ziying Tan, Han Yu 2/24/2026

Federated Learning Playground

Interactive browser-based educational platform for learning Federated Learning concepts with real-time visualization of heterogeneous data effects.

Ax Pascal Jr Tikeng Notsawo, Guillaume Dumas, Guillaume Rabusseau 2/24/2026

Grokking Finite-Dimensional Algebra

Investigation of grokking phenomenon in neural networks learning multiplication in finite-dimensional algebras beyond group operations.