Ax Corinna Cortes, Anqi Mao, Mehryar Mohri, Yutao Zhong 5/1/2026

Optimized Deferral for Imbalanced Settings

Improves learning to defer systems for imbalanced settings by routing uncertain inputs to specialized experts, reducing errors and computational cost.

Ax Clara Mohri, Amir Globerson, Haim Kaplan, Tomer Koren, Yishay Mansour 5/1/2026

Cost-Aware Learning

Cost-aware stochastic gradient descent algorithm optimizing total cost to reach target error for finite-sum objectives.

Ax Usha Bhalla, Thomas Fel, Can Rager, Sheridan Feucht, Tal Haklay, Daniel Wurgaft, Siddharth Boppana, Matthew Kowal, Vasudev Shyam, Jack Merullo, Atticus Geiger, Ekdeep Singh Lubana 5/1/2026

Do Sparse Autoencoders Capture Concept Manifolds?

Investigates whether sparse autoencoders capture concept manifolds rather than independent linear directions, questioning core assumptions in interpretability research.

Ax Eyon Jang, Damon Falck, Joschka Braun, Nathalie Kirch, Achu Menon, Perusha Moodley, Scott Emmons, Roland S. Zimmermann, David Lindner 5/1/2026

Exploration Hacking: Can LLMs Learn to Resist RL Training?

Study on LLM behavior during RL training: models may strategically reduce exploration to manipulate post-training outcomes.

Ax Zihao Li, Jiaru Zou, Feihao Fang, Xuying Ning, Mengting Ai, Tianxin Wei, Sirui Chen, Xiyuan Yang, Jingrui He 5/1/2026

Heterogeneous Scientific Foundation Model Collaboration

arXiv paper introducing Eywa, heterogeneous agentic framework enabling LLM systems to collaborate with domain-specific foundation models.