Ax Arthur Chen, Zuxin Liu, Jianguo Zhang, Akshara Prabhakar, Zhiwei Liu, Shelby Heinecke, Silvio Savarese, Victor Zhong, Caiming Xiong 2/24/2026

Test-Time Adaptation for LLM Agents via Environment Interaction

Method for adapting LLM agents to novel environments through test-time interaction, addressing syntactic and semantic mismatches in observation formats and state dynamics.

Ax Xiao Wu, Ting-Zhu Huang, Liang-Jian Deng, Xiaobing Yu, Yu Zhong, Shangqi Deng, Ufaq Khan, Jianghao Wu, Xiaofeng Liu, Imran Razzak, Xiaojun Chang, Yutong Xie 2/24/2026

SelfAI: A self-directed framework for long-horizon scientific discovery

SelfAI multi-agent system for self-directed long-horizon scientific discovery with human-in-the-loop workflows and exploration trade-offs.

Ax Yaswanth Chittepu, Raghavendra Addanki, Tung Mai, Anup Rao, Branislav Kveton 2/24/2026

ML-Tool-Bench: Tool-Augmented Planning for ML Tasks

ML-Tool-Bench framework for tool-augmented planning in autonomous ML agents orchestrating data analysis and model optimization workflows.

Ax Yifan Zhang, Zixiang Chen, Yifeng Liu, Zhen Qin, Huizhuo Yuan, Kangping Xu, Yang Yuan, Quanquan Gu, Andrew Chi-Chih Yao 2/24/2026

Group Representational Position Encoding

GRAPE framework unifying positional encoding mechanisms using group actions for multiplicative rotations and additive biases.

Ax Kecheng Cai, Chao Peng, Chenyang Xu, Xia Chen, Yi Wang, Shuo Shi, Qiyuan Liang 2/24/2026

Self-Augmented Mixture-of-Experts for QoS Prediction

Mixture-of-experts model with self-augmentation for Quality of Service prediction in web service recommendation systems.

HN seawolf2357 2/24/2026

Do Bubbles Form When AIs Simulate Capitalism?

Research paper demonstrating multiple AI agents connected to live trading APIs all bankrupted within 30 minutes due to LLM hallucination causing false market citations.

HN ramoz 2/24/2026

Get to Know OpenClaw Security

Security guide for OpenClaw, a self-hosted AI agent gateway connecting LLMs to messaging platforms (Slack, Discord, Telegram) with tool access and local execution.

HN fagnerbrack 2/24/2026

Getting Real with LLMs

Critical analysis mapping real-world engineering tasks to LLM/agent tool capabilities, distinguishing genuine functionality from hype in ecosystem claims.