Ax Tianyi Jiang, Arctanx An, Hengyi Feng, Naixin Zhai, Haodong Li, Xiaomin Yu, Jiahui Liu, Hanwen Du, Shuo Zhang, Zhi Yang, Jie Huang, Youhua Li, Yongxin Ni, Huacan Wang, Ronghao Chen 3/19/2026

Chain of Mindset: Reasoning with Adaptive Cognitive Modes

Chain of Mindset method enabling LLMs to adaptively switch between cognitive modes for improved reasoning across problem-solving stages.

Ax Cheng Zhen, Prayoga, Nischal Aryal, Arash Termehchy, Garrett Biwer, Lubna Alzamil 3/19/2026

Learning Over Dirty Data with Minimal Repairs

Research on minimal data repair showing imputing all missing values unnecessary for accurate ML models; introduces minimal and almost-minimal repair concepts.

Ax Mathew J. Koretsky, Maya Willey, Owen Bianchi, Chelsea X. Alvarado, Tanay Nayak, Nicole Kuznetsov, Sungwon Kim, Mike A. Nalls, Daniel Khashabi, Faraz Faghri 3/19/2026

BiomedSQL: Text-to-SQL for Scientific Reasoning on Biomedical Knowledge Bases

Benchmark for text-to-SQL systems evaluating scientific reasoning over biomedical knowledge bases requiring implicit domain understanding.

Ax Yuxiang Ji, Ziyu Ma, Yong Wang, Guanhua Chen, Xiangxiang Chu, Liaoni Wu 3/19/2026

Tree Search for LLM Agent Reinforcement Learning

Tree-based group relative policy optimization for LLM agent reinforcement learning addressing sparse supervision in long-horizon multi-turn tasks.

Ax Andy Dimnaku, Abdullah Yusuf Kavranoglu, Yaser Abu-Mostafa 3/19/2026

Generative Hints

Generative hints training methodology enforcing functional invariances in vision models beyond empirical training data.

Ax Nathan Breslow, Aayush Mishra, Mahler Revsine, Michael C. Schatz, Anqi Liu, Daniel Khashabi 3/19/2026

Genomic Next-Token Predictors are In-Context Learners

Research showing genomic sequence models exhibit in-context learning similar to LLMs, demonstrating ICL emerges across sequence domains.

Ax Di Feng, Kaixin Ma, Feng Nan, Haofeng Chen, Bohan Zhai, David Griffiths, Mingfei Gao, Zhe Gan, Eshan Verma, Yinfei Yang, Zhifeng Chen, Afshin Dehghan 3/19/2026

SO-Bench: A Structural Output Evaluation of Multimodal LLMs

Benchmark for evaluating multimodal LLMs on schema-grounded visual information extraction and reasoning tasks in agentic settings.