Ax Nicolas M\"uller, Piotr Kawa, Adriana Stan, Thien-Phuc Doan, Souhwan Jung, Wei Herng Choong, Philip Sperl, Konstantin B\"ottinger 5/14/2026

DeePen: Penetration Testing for Audio Deepfake Detection

DeePen: penetration testing methodology for evaluating robustness of machine learning audio deepfake detection classifiers.

Ax Zhenhe Wu, Jian Yang, Zhongjiang He, Changzai Pan, Jie Zhang, Jiaheng Liu, Xianjie Wu, Yu Zhao, Shuangyong Song, Yongxiang Li, Zhoujun Li, Xueling Li 5/14/2026

Table-R1: Region-based Reinforcement Learning for Table Understanding

Region-based reinforcement learning approach optimizing LLM performance for table question-answering with structured row-column reasoning.

Ax Jeffrey T. H. Wong, Cheng Zhang, Xinye Cao, Pedro Gimenes, Christos-Savvas Bouganis, George A. Constantinides, Wayne Luk, Yiren Zhao 5/14/2026

A3 : an Analytical Low-Rank Approximation Framework for Attention

A3: analytical low-rank approximation framework for transformer attention mechanisms to enable efficient LLM compression and deployment.

Ax Jehyeok Yeon, Isha Chaudhary, Gagandeep Singh 5/14/2026

Quantitative Certification of Agentic Tool Selection

Framework for certifying tool selection in LLM-based agentic systems, evaluating robustness against adversarial tool pools and deployment scenarios.

Ax Luca Belli, Kate H. Bentley, Will Alexander, Emily Ward, Matt Hawrilenko, Kelly Johnston, Mill Brown, Adam M. Chekroud 5/14/2026

VERA-MH Concept Paper

Automated safety evaluation framework for mental health AI chatbots using clinician-informed rubric and multi-agent validation approach.

Ax Dario Shariatian, Alain Durmus, Umut Simsekli, Stefano Peluchetti 5/14/2026

Latent-Augmented Discrete Diffusion Models

Latent-augmented discrete diffusion model with auxiliary channel for improved few-step language generation and cross-token dependencies.

Ax John Yang, Kilian Lieret, Joyce Yang, Carlos E. Jimenez, Muhtasham Oblokulov, Aryan Siddiqui, Ofir Press, Ludwig Schmidt, Diyi Yang 5/14/2026

CodeClash: Benchmarking Goal-Oriented Software Engineering

Benchmark evaluating whether LLMs can iteratively develop code toward high-level goals beyond isolated task completion.

Ax Yifan Zhang, Yifeng Liu, Mengdi Wang, Quanquan Gu 5/14/2026

Deep Delta Learning

Deep Delta Learning introduces residual update rule for transformer layers enabling selective content rewriting in deep networks.

Ax Or Ordentlich, Yury Polyanskiy 5/14/2026

High-Rate Quantized Matrix Multiplication I

Research on quantized matrix multiplication optimization for efficient LLM deployment with weight and activation quantization.

Ax Tomas Ruiz, Zhen Qin, Yifan Zhang, Xuyang Shen, Yiran Zhong, Mengdi Wang 5/14/2026

FlashSampling: Fast and Memory-Efficient Exact Sampling

Optimized sampling primitive that fuses categorical sampling into LM-head computation, reducing memory traffic for large-vocabulary decoding.

Ax Kai-Wei Chang, Wei-Chih Chen, En-Pei Hu, Hung-yi Lee, James Glass 5/14/2026

TiCo: Time-Controllable Spoken Dialogue Model

Spoken dialogue model with controllable response duration for voice assistants and interactive agents.