Ax Drew Prinster, Clara Fannjiang, Ji Won Park, Kyunghyun Cho, Anqi Liu, Suchi Saria, Samuel Stanton 4/17/2026

Conformal Policy Control

Conformal inference method to regulate untested policies against safe reference policies for high-stakes agent deployment.

Ax Richard Servajean, Philippe Servajean 4/17/2026

Measuring the metacognition of AI

Develops methods to measure metacognitive capabilities of AI systems for uncertainty assessment and decision reliability.

Ax Myeongsoo Kim, Joe Hsu, Dingmin Wang, Shweta Garg, Varun Kumar, Murali Krishna Ramanathan 4/17/2026

CODESTRUCT: Code Agents over Structured Action Spaces

CODESTRUCT reframes code editing from text manipulation to structured AST operations, enabling LLM code agents to apply syntax-validated transformations on repositories.

Ax Georgianna "Blue" Lin, Rencong Jiang, No\'emie Elhadad, Xuhai "Orson" Xu 4/17/2026

A longitudinal health agent framework

Proposes longitudinal health agent framework addressing user intent and accountability for multi-turn health tasks like symptom management and behavior change.

Ax Serdar Kadioglu, Karthik Uppuluri, Akash Singirikonda 4/17/2026

Modeling Copilots for Text-to-Model Translation

Text2Model suite introduces LLM-based copilots for text-to-optimization translation with varying complexity strategies and cross-domain dataset Text2Zinc.

Ax Yuwei Yin, Giuseppe Carenini 4/17/2026

Improving Language Models with Intentional Analysis

Introduces intentional analysis framework to improve language model reasoning by explicitly modeling intent as cognitive notion underlying human communication and problem-solving.

Ax Yiyuan Yang, Zichuan Liu, Lei Song, Kai Ying, Zhiguang Wang, Tom Bamford, Svitlana Vyetrenko, Jiang Bian, Qingsong Wen 4/17/2026

Time-RA: Towards Time Series Reasoning for Anomaly Diagnosis with LLM Feedback

Time-RA reformulates time series anomaly detection as a generative reasoning task using LLM feedback, introducing RATs40K dataset for fine-grained anomaly categorization and explanation.

Ax Jiahao Tang, Henry Hengyuan Zhao, Lijian Wu, Yifei Tao, Dongxing Mao, Yang Wan, Jingru Tan, Min Zeng, Min Li, Alex Jinpeng Wang 4/17/2026

From Charts to Code: A Hierarchical Benchmark for Multimodal Models

Chart2Code benchmark evaluating chart understanding and code generation in multimodal LLMs across three difficulty levels with real-world scenarios.