Ax Jichao Wang, Liuyang Bian, Yufeng Zhou, Han Xiao, Yue Pan, Guozhi Wang, Hao Wang, Zhaoxiong Wang, Yafei Wen, Xiaoxin Chen, Shuai Ren, Lingfang Zeng 4/27/2026

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning

SOLAR-RL: Semi-online reinforcement learning framework for training multimodal LLM-based GUI agents on complex navigation tasks.

Ax Bilal Faye, Abdoulaye Mbaye, Hanane Azzag, Mustapha Lebbah 4/27/2026

Adaptive Head Budgeting for Efficient Multi-Head Attention

Proposes adaptive head budgeting mechanism for multi-head attention in Transformers to improve efficiency by selectively activating attention heads based on task requirements.

Ax Akram Erraqabi, Michal Valko, Alexandra Carpentier, Odalric-Ambrym Maillard 4/27/2026

Pliable rejection sampling

ArXiv paper on pliable rejection sampling with kernel estimators for sampling difficult distributions. Machine learning research with new approach.

Ax Zhanli Wu, Fabrizio Leisen, Miguel-Angel Luque-Fernandez, F. Javier Rubio 4/27/2026

Conformalized Super Learner

ArXiv paper on Conformalized Super Learner ensemble method with interval prediction uncertainty quantification. Machine learning research with theoretical contributions.