Ax Hieu Bui, Ziyan Gao, Yuya Hosoda, Joo-Ho Lee 2/24/2026

Visual Prompt Guided Unified Pushing Policy

Unified pushing policy for robotic manipulation using visual prompting with broader applicability across scenarios.

Ax Shirui Chen, Cole Harrison, Ying-Chun Lee, Angela Jin Yang, Zhongzheng Ren, Lillian J. Ratliff, Jiafei Duan, Dieter Fox, Ranjay Krishna 2/24/2026

TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics

TOPReward uses token probabilities from VLA models as zero-shot reward signals for robot RL training with improved sample efficiency.

Ax Zelin He, Boran Han, Xiyuan Zhang, Shuai Zhang, Haotian Lin, Qi Zhu, Haoyang Fang, Danielle C. Maddix, Abdul Fatir Ansari, Akash Chandrayan, Abhinav Pradhan, Bernie Wang, Matthew Reimherr 2/24/2026

SenTSR-Bench: Thinking with Injected Knowledge for Time-Series Reasoning

Benchmark for time-series reasoning combining general LLM reasoning with domain-specific time-series knowledge through injected knowledge approach.

Ax Bryan Guanrong Shan, Alysa Ziying Tan, Han Yu 2/24/2026

Federated Learning Playground

Interactive browser-based platform for learning federated learning concepts with real-time visualization of heterogeneous data distributions and aggregation algorithms.

Ax Arundhati Banerjee, Jeff Schneider 2/24/2026

Cost-Aware Diffusion Active Search

Active search algorithm for autonomous agents balancing exploration-exploitation with cost-aware decision making.

Ax Girmaw Abebe Tadesse, Titien Bartette, Andrew Hassanali, Allen Kim, Jonathan Chemla, Andrew Zolli, Yves Ubelmann, Caleb Robinson, Inbal Becker-Reshef, Juan Lavista Ferres 2/24/2026

Satellite-Based Detection of Looted Archaeological Sites Using Machine Learning

Machine learning pipeline using satellite imagery to detect looted archaeological sites in Afghanistan. Application of computer vision.

Ax Jesse Farebrother, Matteo Pirotta, Andrea Tirinzoni, Marc G. Bellemare, Alessandro Lazaric, Ahmed Touati 2/24/2026

Compositional Planning with Jumpy World Models

Compositional planning approach using pre-trained policy composition for temporal abstraction in agent decision-making.

Ax Mohammed Javed Absar, Muthu Baskaran, Abhikrant Sharma, Abhilash Bhandari, Ankit Aggarwal, Arun Rangasamy, Dibyendu Das, Fateme Hosseini, Franck Slama, Iulian Brumar, Jyotsna Verma, Krishnaprasad Bindumadhavan, Mitesh Kothari, Mohit Gupta, Ravishankar Kolachana, Richard Lethin, Samarth Narang, Sanjay Motilal Ladwa, Shalini Jain, Snigdha Suresh Dalvi, Tasmia Rahman, Venkat Rasagna Reddy Komatireddy, Vivek Vasudevbhai Pandya, Xiyue Shi, Zachary Zipper 2/24/2026

Hexagon-MLIR: An AI Compilation Stack For Qualcomm's Neural Processing Units (NPUs)

Open-source MLIR-based compilation stack for deploying Triton kernels and PyTorch models on Qualcomm NPUs.