Ax Shi Feng, Hanlin Zhang, Fan Nie, Sham Kakade, Yiling Chen 7/10/2026

Peer-Predictive Self-Training for Language Model Reasoning

Peer-Predictive Self-Training (PST) framework enables collaborative self-improvement of language models without external supervision using cross-model aggregation.

Ax Bingxi Zhao, Jiahao Zhang, Xubin Ren, Zirui Guo, Tianzhe Chu, Yi Ma, Chao Huang 7/10/2026

DeepTutor: Towards Agentic Personalized Tutoring

Open-source agentic tutoring framework combining LLMs with personalized feedback, difficulty calibration, and citation-grounded problem solving.

Ax Liang Luo, Yinbin Ma, Quanyu Zhu, Vasiliy Kuznetsov, Yuxin Chen, Neng Shi, Jian Jiao, Jiecao Yu, Buyun Zhang, Tongyi Tang, Xiaohan Wei, Yanli Zhao, Zeliang Chen, Yuchen Hao, Venkatesh Ranganathan, Sandeep Parab, Yantao Yao, Maxim Naumov, Chunzhi Yang, Shen Li, Ellie Wen, Wenlin Chen, Santanu Kolay, Chunqiang Tang 7/10/2026

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale

Low-precision FP8 arithmetic optimization techniques for large recommendation models, addressing numerical sensitivity and training efficiency.

Ax Ian Rios-Sialer, Shantanu Darveshi, Shuai Jiang, Avigya Paudel, Anastasiia Pronina, Ipshita Bandyopadhyay, Justin Shenk 7/10/2026

Temporal Preference Concepts and their Functions in a Large Language Model

arXiv paper identifying and analyzing temporal preference subgraphs in LLMs through causal localization, showing how models represent temporal tradeoffs internally.

Ax Abhinav Agarwal, Adam Wei, Taylan Kargin, Michael Zeng, Cole Becker, Arif Kerem Dayi, Pablo Parrilo, Asuman Ozdaglar, Russ Tedrake 7/10/2026

Training and Evaluating Diffusion Policies with Long Context Lengths

Benchmark of diffusion policies for robotic manipulation with incrementally increasing context lengths to enable memory and long-horizon task performance.

Ax Jannik H\"osch, Alessandro Sestini, Florian Fuchs, Amir Baghi, Joakim Bergdahl, Iolanda Leite, Konrad Tollmar, Jean-Philippe Barrette-LaPierre, Linus Gissl\'en 7/10/2026

Hierarchical Control in Multi-Agent Games: LLM-based Planning and RL Execution

Hierarchical multi-agent architecture combining LLM-based strategic planning with specialized RL skill policies for complex coordinated decision-making.

Ax Yanis Labrak, Dairazalia Sanchez-Cortes, Sergio Burdisso, S\'everin Baroudi, Shashi Kumar, Esa\'u Villatoro-Tello, Srikanth Madikeri, Manjunath K E, Old\v{r}ich Plchot, Kadri Hacio\u{g}lu, Petr Motlicek, Andreas Stolcke 7/10/2026

How to Leverage Synthetic Speech for LLM-Based ASR Systems?

Study on leveraging synthetic TTS speech for training ASR systems in privacy-constrained domains like banking and healthcare, addressing synthetic-real data gaps.

Ax R\'ois\'in Luo, Christian Gagn\'e, Jonas Ngnaw\'e, Ihsan Ullah, Karyn Morrissey 7/10/2026

A Stochastic--Geometric Theory of Scaling Laws in Grokking

Theoretical characterization of grokking phenomenon using stochastic-geometric analysis of solution space topology and delayed generalization in neural networks.

Ax Zhuoxuan Zhang (Yang), Kangqi Ni (Yang), Yuhang Chen (Yang), Mingfu Liang (Yang), Xiaohan Wei (Yang), Yunchen Pu (Yang), Fei Tian (Yang), Chonglin Sun (Yang), Frank Shyu (Yang), Adam (Yang), Song, Sandeep Pandey, Luke Simon, Tianlong Chen, Xi Liu 7/10/2026

Diffusion-GR2: Diffusion Generative Reasoning Re-ranker

Diffusion-GR2 uses block-diffusion language models for faster generative reasoning re-ranking with parallel decoding instead of sequential autoregressive inference.

Ax Vaishnavi Sinha, Pooja Guttal, Pranay Deep Reddy Katike, Vishal Sinha, Gerald Ndawula, Lira Yoon, Andrea Kleinsmith, Manas Gaur 7/10/2026

Where do LLMs Fall Short in CBT-Guided Affective Reasoning?

Study analyzing LLM failures in applying Cognitive Behavioral Therapy frameworks despite high theoretical knowledge, highlighting reasoning gaps in practical application.

Ax Alexis Kafantaris 7/10/2026

LLM for the development of FCM

Approach using local LLMs to extract quantitative data from text for developing fuzzy cognitive maps.