Ax Xingze Zou, Jing Wang, Yuhua Zheng, Xueyi Chen, Haolei Bai, Lingcheng Kong, Syed A. R. Abu-Bakar, Zhaode Wang, Chengfei Lv, Haoji Hu, Huan Wang 3/17/2026

MobileKernelBench: Can LLMs Write Efficient Kernels for Mobile Devices?

MobileKernelBench benchmark evaluating LLM capability to generate efficient computational kernels for mobile devices, with systematic investigation of code generation limits.

Ax Zheda Mai, Ke Zhang, Fu-En Wang, Zixiao Ken Wang, Albert Y. C. Chen, Lu Xia, Min Sun, Wei-Lun Chao, Cheng-Hao Kuo 3/17/2026

Revisiting Model Stitching In the Foundation Model Era

Research on model stitching technique for Vision Foundation Models, testing representational compatibility across models with different training objectives and data sources.

Ax Dayuan Fu, Shenyu Wu, Yunze Wu, Zerui Peng, Yaxing Huang, Jie Sun, Ji Zeng, Mohan Jiang, Lin Zhang, Yukun Li, Jiarui Hu, Liming Liu, Jinlong Hou, Pengfei Liu 3/17/2026

daVinci-Env: Open SWE Environment Synthesis at Scale

Large-scale open-source software engineering environment for training AI agents with executable, verifiable tasks and dynamic feedback.

Ax Daniel Bretsko, Piotr Walas, Devashish Khulbe, Sebastian Stros, Stanislav Sobolevsky, Tomas Satura 3/17/2026

FastODT: A tree-based framework for efficient continual learning

Tree-based continual learning framework for non-stationary data distributions with constrained computational resources in time series applications.

Ax Thibault Formal, Maxime Louis, Herv\'e Dejean, St\'ephane Clinchant 3/17/2026

Learning Retrieval Models with Sparse Autoencoders

Sparse autoencoders foundation for learned sparse retrieval, decomposing LLM representations into interpretable latent features for efficient document retrieval.

Ax Sunghyeon Woo, Jaeeun Kil, Hoseung Kim, Minsub Kim, Joonghoon Kim, Ahreum Seo, Sungjae Lee, Minjung Jo, Jiwon Ryu, Baeseong Park, Se Jung Kwon, Dongsoo Lee 3/17/2026

ICaRus: Identical Cache Reuse for Efficient Multi Model Inference

Multi-model inference optimization reusing identical KV caches across models to reduce memory consumption in agentic AI systems.

Ax Angelika Romanou, Mark Ibrahim, Candace Ross, Chantal Shaib, Kerem Okta, Sam Bell, Elia Ovalle, Jesse Dodge, Antoine Bosselut, Koustuv Sinha, Adina Williams 3/17/2026

Brittlebench: Quantifying LLM robustness via prompt sensitivity

Benchmark quantifying LLM robustness by measuring model sensitivity to prompt variations, typos, and paraphrases in real-world conditions.

Ax Xinrun Xu, Pi Bu, Ye Wang, B\"orje F. Karlsson, Ziming Wang, Tengtao Song, Qi Zhu, Jun Song, Shuo Zhang, Zhiming Ding, Bo Zheng 3/17/2026

ICPRL: Acquiring Physical Intuition from Interactive Control

ICPRL framework enabling vision language models to learn physical reasoning from pixel-based interactive control.

Ax Nirmalendu Prakash, Narmeen Oozeer, Michael Lan, Luka Samkharadze, Phillip Howard, Roy Ka-Wei Lee, Dhruv Nathawani, Shivam Raval, Amirali Abdullah 3/17/2026

DreamReader: An Interpretability Toolkit for Text-to-Image Models

DreamReader: unified interpretability toolkit for analyzing text-to-image diffusion models with causal and representation analysis.

Ax Wei-Hao Wu, Ting-Zhu Huang, Xi-Le Zhao, Yisi Luo, Deyu Meng 3/17/2026

Neural Approximation and Its Applications

Neural basis functions using untrained networks for multivariate function approximation in machine learning.

Ax Florin Leon 3/17/2026

Modular Neural Computer

Modular Neural Computer is a memory-augmented architecture combining external associative memory with functional MLP modules for exact algorithmic computation.