Ax Lily Hong Zhang, Smitha Milli, Karen Jusko, Jonathan Smith, Brandon Amos, Wassim Bouaziz, Manon Revel, Jack Kussman, Yasha Sheynin, Lisa Titus, Bhaktipriya Radharapu, Jane Yu, Vidya Sarma, Kris Rose, Maximilian Nickel 2/23/2026

Cultivating Pluralism In Algorithmic Monoculture: The Community Alignment Dataset

Large-scale multilingual study (N=15,000) showing humans have more preference variation than LLMs; introduces Community Alignment Dataset for cultural pluralism.

Ax Yuehan Qin, Li Li, Defu Cao, Tiankai Yang, Jiate Li, Yue Zhao 2/23/2026

M3OOD: Automatic Selection of Multimodal OOD Detectors

Automatic selection method for multimodal out-of-distribution detectors addressing robustness across diverse distribution shifts in video, audio, and sensor data.

Ax Michael Sullivan, Alexander Koller 2/23/2026

GRPO is Secretly a Process Reward Model

Theoretical proof that GRPO RL algorithm with outcome reward models is equivalent to process reward models with Monte-Carlo-based weighting.

Ax Hoang Phan, Sungmin Cha, Tung Lam Tran, Qi Lei 2/23/2026

Toward a Holistic Approach to Continual Model Merging

Framework for continual model merging addressing scalability of task vectors and functional drift through pre-, during-, and post-merging interventions.

Ax Nimrod Berman, Assaf Hallak, Assaf Shocher 2/23/2026

Who Said Neural Networks Aren't Linear?

Theoretical work identifying non-standard vector spaces where neural networks act as linear operators using transport of structure from algebra.

Ax Minseo Kim, Chenfeng Xu, Coleman Hooper, Harman Singh, Ben Athiwaratkun, Ce Zhang, Kurt Keutzer, Amir Gholami 2/23/2026

CDLM: Consistency Diffusion Language Models For Faster Sampling

CDLM accelerates diffusion language models through consistency modeling and enables KV caching for faster parallel generation.

Ax Julian Kleutgens, Claudio Battiloro, Lingkai Kong, Benjamin Grewe, Francesca Dominici, Mauricio Tec 2/23/2026

Guided Transfer Learning for Discrete Diffusion Models

Transfer learning approach for discrete diffusion models in small-data regimes using classifier ratio-based guidance adapted from continuous models.

Ax Derrick Gilchrist Edward Manoharan, Anubha Goel, Alexandros Iosifidis, Henri Hansen, Juho Kanniainen 2/23/2026

Learning hidden cascades via classification

Addresses learning spreading dynamics in social networks with hidden individual statuses using classification methods with observable intermediate indicators.

Ax Samuele Marro, Jialin Yu, Emanuele La Malfa, Oishi Deb, Jiawei Li, Yibo Yang, Ebey Abraham, Sunando Sengupta, Eric Sommerlade, Michael Wooldridge, Philip Torr 2/23/2026

Benchmarking at the Edge of Comprehension

Analysis of LLM benchmark saturation and the challenge of creating discriminative tasks as frontier models improve, discussing feasibility of future benchmarking.