Ax Jingyu Zhang, Tianjian Li, William Jurayj, Hongyuan Zhan, Benjamin Van Durme, Daniel Khashabi 4/13/2026

Many-Tier Instruction Hierarchy in LLM Agents

Instruction Hierarchy in LLM Agents arXiv paper addressing multi-source conflicting instructions in LLM systems. Examines privilege levels for safe instruction following.

Ax Maksim Anisimov (Imperial College London), Francesco Belardinelli (Imperial College London), Matthew Wicker (Imperial College London) 4/13/2026

SafeAdapt: Provably Safe Policy Updates in Deep Reinforcement Learning

SafeAdapt arXiv paper on provably safe policy updates in deep RL for non-stationary environments. Addresses safety preservation during policy changes.

Ax Stefan Andreas Baumann, Jannik Wiese, Tommaso Martorella, Mahdi M. Kalayeh, Bj\"orn Ommer 4/13/2026

Envisioning the Future, One Step at a Time

Method for predicting future scene evolution by modeling uncertainty and simulating trajectories rather than dense pixel-level changes.

Ax Xiaojie Xu, Zongyuan Li, Chang Lu, Runnan Qi, Yanan Ni, Lumin Jiang, Xiangbei Liu, Xuebo Zhang, Yongchun Fang, Kuihua Huang, Xian Guo, Zhanghua Wu, Zhenya Li 4/13/2026

Reflection of Episodes: Learning to Play Game from Expert and Self Experiences

Framework enabling LLMs to learn complex game strategies through self-reflection on expert and self-generated experiences in StarCraft II.

Ax Shahab Rahimirad, Guven Gergerli, Lucia Romero, Angela Qian, Matthew Lyle Olson, Simon Stepputtis, Joseph Campbell 4/13/2026

Bayesian Social Deduction with Graph-Informed Language Models

Study evaluating LLM performance on social reasoning tasks in the Avalon game, testing inference capabilities and model distillation effects.

Ax Zhenfeng Lin, Haoji Hu, Ming Hao, Xuchao Zhang, Ryan Zhang, Junhao Li, Ze Li, Oleg Kulygin, Chetan Bansal, Hatay Tuna, Murali Chintalapati, Sheila Jiang, Salman Zafar, Angie Anderson 4/13/2026

ActionNex: A Virtual Outage Manager for Cloud Computing

Production agentic system for cloud outage management with real-time updates, knowledge distillation, and conditioned action recommendations.

Ax Jingyang Qiao, Weicheng Meng, Yu Cheng, Zhihang Lin, Zhizhong Zhang, Xin Tan, Jingyu Gong, Kun Shao, Yuan Xie 4/13/2026

Memory Intelligence Agent

Memory system for deep research agents enabling efficient evolution and reasoning through intelligent trajectory memory management.

Ax Wenxuan Liu, Zixuan Li, Long Bai, Chunmao Zhang, Fenghui Zhang, Zhuo Chen, Wei Li, Yuxin Zuo, Fei Wang, Bingbing Xu, Xuhui Jiang, Jin Zhang, Xiaolong Jin, Jiafeng Guo, Tat-Seng Chua, Xueqi Cheng 4/13/2026

Towards Knowledgeable Deep Research: Framework and Benchmark

Framework and benchmark for deep research agents using structured knowledge alongside unstructured web content for comprehensive reports.