Ax Han Yi Shin, Heeju Ko, Jaewon Mun, Qixing Huang, Jaehyeok Lee, Sung June Kim, Honglak Lee, Sujin Jang, Sangpil Kim 5/14/2026

SECOND-Grasp: Semantic Contact-guided Dexterous Grasping

Approach integrating dexterous grasping with language-guided semantics for reliable robotic manipulation.

Ax Xinyu Liu, Kechen Jiao, Chunyang Xiao, Runsong Zhao, Junhao Ruan, Bei Li, Jiahao Liu, Qifan Wang, Xin Chen, Jingang Wang, Tong Xiao, JingBo Zhu 5/14/2026

Teacher-Guided Policy Optimization for LLM Distillation

LLM distillation method using teacher guidance and Reverse KL to improve student model learning from divergent distributions.

Ax Ian Osband 5/14/2026

Delightful Exploration

Exploration algorithm balancing uncertainty reduction with expected improvement in large action spaces.

Ax Yunheng Wang, Yuetong Fang, Taowen Wang, Lusong Li, Kun Liu, Junzhe Xu, Zizhao Yuan, Yixiao Feng, Jiaxi Zhang, Wei Lu, Zecui Zeng, Renjing Xu 5/14/2026

What Limits Vision-and-Language Navigation ?

Analysis of embodied AI agents' limitations in vision-language navigation when transitioning from simulation to real-world deployment.

Ax Viktor Moskvoretskii, Dominik Glandorf, Jorge Medina Moreira, Tanja K\"aser, Robert West 5/14/2026

Tracing Persona Vectors Through LLM Pretraining

Interpretability research tracing how LLMs represent behavioral traits like sycophancy as linear directions in internal activations.

Ax Katarzyna Kobalczyk, Mihaela van der Schaar 5/14/2026

Discovery of Hidden Miscalibration Regimes

Method for discovering localized model calibration failures beyond global reliability metrics using structured analysis.

Ax Asim Osman, Sasha Abramowitz, Mark Bergh, Ulrich Armel Mbou Sob, Ruan John de Kock, Omayma Mahjoub, Oussama Hidaoui, Noah De Nicola, Arnol Manuel Fokam, Felix Chalumeau, Daniel Rajaonarivonivelomanantsoa, Siddarth Singh, Refiloe Shabe, Juan Claude Formanek, Simon Verster Du Toit, Arnu Pretorius 5/14/2026

Self-Supervised On-Policy Reinforcement Learning via Contrastive Proximal Policy Optimisation

arXiv paper on AttenA+: improving robotic foundation models by addressing temporal heterogeneity in manipulation tasks.

Ax Steven Seiden, Triss Ren, Caroline Zhang, Taein Kim, Enze Liu, Emily Wenger 5/14/2026

Identifying AI Web Scrapers Using Canary Tokens

Method using canary tokens to detect and identify web scraping by AI systems for LLM training data collection.

Ax Qian Shen (University of Florida, Gainesville, USA), Fanghua Cao (University of Florida, Gainesville, USA), Min Yao (University of Florida, Gainesville, USA), Shlok Gilda (University of Florida, Gainesville, USA), Bonnie J. Dorr (University of Florida, Gainesville, USA), Walter L. Leite (University of Florida, Gainesville, USA) 5/14/2026

Children's English Reading Story Generation via Supervised Fine-Tuning of Compact LLMs with Controllable Difficulty and Safety

Fine-tuned compact LLMs generate children's reading stories with controllable difficulty and safety constraints.

Ax Mind Lab, :, Song Cao, Vic Cao, Andrew Chen, Kaijie Chen, Cleon Cheng, Steven Chiang, Kaixuan Fan, Hera Feng, Huan Feng, Arthur Fu, Jun Gao, Hongquan Gu, Aaron Guan, Nolan Ho, Mutian Hong, Hailee Hou, Peixuan Hua, Charles Huang, Miles Jiang, Nora Jiang, Yuyi Jiang, Qiuyu Jin, Fancy Kong, Andrew Lei, Kyrie Lei, Alexy Li, Lucian Li, Ray Li, Theo Li, Zhihui Li, Jiayi Lin, Kairus Liu, Kieran Liu, Logan Liu, Xiang Liu, Irvine Lu, Maeve Luo, Runze Lv, Pony Ma, Verity Niu, Anson Qiu, Vincent Wang, Rio Yang, Maxwell Yao, Carrie Ye, Regis Ye, Wenlin Ye, Josh Ying, Danney Zeng, Yuhan Zhan, Anya Zhang, Di Zhang, Ruijia Zhang, Sueky Zhang, Ya Zhang, Wei Zhao, Ada Zhou, Changhai Zhou, Yuhua Zhou, Xinyue Zhu, Murphy Zhuang 5/14/2026

MinT: Managed Infrastructure for Training and Serving Millions of LLMs

MinT infrastructure system for efficient LoRA fine-tuning and serving of millions of LLM variants using shared base models.