Ax Donato Crisostomi 5/6/2026

Model Merging: Foundations and Algorithms

Thesis on model merging paradigm: combining independently trained neural networks in weight space without optimization or original training data access.

Ax Rohit Agarwal, Joshua Lin, Mark Braverman, Elad Hazan 5/6/2026

AI Alignment via Incentives and Correction

Studies AI alignment through law-and-economics models of deterrence, analyzing how agentic AI systems respond strategically to incentive structures.

Ax Haoshen Zhang, Di Wen, Kunyu Peng, David Schneider, Zeyun Zhong, Alexander Jaus, Zdravko Marinov, Jiale Wei, Ruiping Liu, Junwei Zheng, Yufan Chen, Yufeng Zhang, Yuanhao Luo, Lei Qi, Rainer Stiefelhagen 5/6/2026

IMPACT-HOI: Supervisory Control for Onset-Anchored Partial HOI Event Construction

IMPACT-HOI mixed-initiative framework for egocentric video annotation of Human-Object Interactions to generate robot manipulation training data.

Ax Hongkun Pan, Yuwei Wu, Wanyi Hong, Shenghui Hu, Qitong Yan, Yi Yang, Rufei Han, Changju Zhou, Minfeng Zhu, Dongming Han, Wei Chen 5/6/2026

Chart-FR1: Visual Focus-Driven Fine-Grained Reasoning on Dense Charts

Chart-FR1: Multimodal LLM benchmark for fine-grained reasoning on high-information-density charts with multiple subplots and dense annotations.

Ax Kyle Lee, Corentin Delacour, Kevin Callahan-Coray, Kyle Jiang, Can Yaras, Samet Oymak, Tathagata Srimani, Kerem Y. Camsari 5/6/2026

Stochastic Sparse Attention for Memory-Bound Inference

SANTA: Sparse attention method for memory-bound LLM inference that samples from post-softmax distribution to reduce KV cache bandwidth at long contexts.

Ax Debeshee Das, Julien Piet, Darya Kaviani, Luca Beurer-Kellner, Florian Tram\`er, David Wagner 5/6/2026

Trojan Hippo: Weaponizing Agent Memory for Data Exfiltration

Trojan Hippo attack exploiting LLM agent memory systems for data exfiltration through dormant payloads planted via single untrusted tool interactions.

Ax Mario Koddenbrock, Christoph Lange, Robin Legner, Martin J\"ager, Martin K\"ogler, Mariano N. Cruz Bournazou, Peter Neubauer, Felix Biessmann, Erik Rodner 5/6/2026

RamanBench: A Large-Scale Benchmark for Machine Learning on Raman Spectroscopy

First large-scale reproducible benchmark for machine learning on Raman spectroscopy, standardizing evaluation across fragmented datasets and spectral models.

Ax Christopher Kelly, Angelica Chowdhury, Alexandra Campili, Bimpe Ayoola, Devin Barbour, Thomas Chen Dawson, Ze Shen Chin, Rokas Gipi\v{s}kis 5/6/2026

Principles and Guidelines for Randomized Controlled Trials in AI Evaluation

Framework for standardizing AI evaluation through randomized controlled trials, adopting validity principles from established experimental disciplines.