Ax Zezhou Zhang, Songxin Zhang, Xiao Xiong, Junjie Zhang, Zejian Xie, Jingyi Xi, Zunyao Mao, Zan Mao, Zhixin Mai, Zhuoyang Song, Jiaxing Zhang 3/16/2026

PVI: Plug-in Visual Injection for Vision-Language-Action Models

Method for injecting auxiliary visual features into vision-language-action models to improve geometric understanding and temporal reasoning for robotic manipulation.

Ax Pierre Moreau, Emeline Pineau Ferrand, Yann Choho, Benjamin Wong, Annabelle Blangero, Milan Bhan 3/16/2026

Towards Faithful Multimodal Concept Bottleneck Models

Research on interpretable multimodal concept bottleneck models ensuring faithful explanations through proper concept detection.

Ax Sibylle Marcotte, Gabriel Peyr\'e, R\'emi Gribonval 3/16/2026

Intrinsic training dynamics of deep neural networks

Theoretical study of implicit bias in deep neural network training showing gradient flow induces learning of lower-dimensional parameter structures.

Ax Giorgos Nikolaou, Tommaso Mencattini, Donato Crisostomi, Andrea Santilli, Yannis Panagakis, Emanuele Rodol\`a 3/16/2026

Language Models are Injective and Hence Invertible

Mathematical proof that transformer language models are injective, enabling exact input recovery from representations despite nonlinear components.

Ax Kemou Li, Qizhou Wang, Yue Wang, Fengpeng Li, Jun Liu, Bo Han, Jiantao Zhou 3/16/2026

LLM Unlearning with LLM Beliefs

Method for unlearning harmful content from LLMs by analyzing belief redistribution in probability space, avoiding unwanted side effects of gradient ascent.