Ax Shaoan Zhao, Huanlin Gao, Qiang Hui, Ting Lu, Xueqiang Guo, Yantao Li, Xinpei Su, Fuyuan Shi, Chao Tan, Fang Zhao, Kai Wang, Shiguo Lian 5/15/2026

MediaClaw: Multimodal Intelligent-Agent Platform Technical Report

MediaClaw: multimodal agent platform with three-layer architecture addressing fragmentation, heterogeneity, and workflow reuse in AIGC deployment.

Ax Netta Madvil, Gilad Dym, Alon Mecilati, Edo Dekel, Jonatan Liberman, Rotem Brazilay, Liron Schliesser, Max Svidlo, Shai Nir, Orel Shalom, Yaron Friedman, David Connack, Amos Rimon, Philip Tannor, Shir Chorev 5/15/2026

Holistic Evaluation and Failure Diagnosis of AI Agents

Holistic evaluation framework for AI agents combining top-down diagnosis with bottom-up span-level analysis to identify failure types and locations.

Ax Baolin Peng, Wenlin Yao, Qianhui Wu, Hao Cheng, Xiao Yu, Rui Yang, Tao Ge, Alessandrio Sordoni, Xingdi Yuan, Yelong Shen, Pengcheng He, Tong Zhang, Zhou Yu, Jianfeng Gao 5/15/2026

Orchard: An Open-Source Agentic Modeling Framework

Orchard: open-source agentic modeling framework for LLM agents with planning, reasoning, tool use, and multi-turn environment interaction.

Ax Shang Zhou, Wenhao Chai, Kaiyuan Liu, Huanzhi Mao, Qiuyang Mang, Jingbo Shang 5/15/2026

OpenDeepThink: Parallel Reasoning via Bradley--Terry Aggregation

OpenDeepThink: test-time scaling method using parallel reasoning with Bradley-Terry aggregation to select best reasoning candidates without ground-truth verification.

Ax Ahmadreza Jeddi, Minh Ngoc Le, Hakki C. Karaimer, Konstantinos G. Derpanis, Babak Taati 5/15/2026

GEAR: Genetic AutoResearch for Agentic Code Evolution

GEAR: genetic algorithm framework enabling autonomous research agents to explore multiple evolutionary paths simultaneously instead of single-path search strategies.

Ax Jiajun Zhou, Wei Shao, Lingchao Zheng, Yuwei Fan, Ngai Wong 5/15/2026

AIS: Adaptive Importance Sampling for Quantized RL

Adaptive importance sampling method for reinforcement learning with quantized rollouts and BF16 trainer mismatch correction.