HN smurda 2/23/2026

Deep-Dive into LLM Fine-Tuning

Guide to fine-tuning LLMs for enterprise applications, covering mechanics of adapting models like Qwen 3 and DeepSeek v3 for domain-specific use cases.

HN IAmSoThirsty 2/23/2026

Mental Health Escalation Router

Production crisis detection system using AI to identify high-risk distress signals with cryptographic audit trails and formal safety guarantees.

Ax Rahul Nanda, Chandra Maddila, Smriti Jha, Euna Mehnaz Khan, Matteo Paltenghi, Satish Chandra 2/23/2026

Wink: Recovering from Misbehaviors in Coding Agents

Wink: Framework for detecting and recovering from misbehaviors in LLM-powered coding agents, addressing issues like instruction deviation, infinite loops, and tool misuse.

Ax Bin Wang, Fan Wang, Pingping Wang, Jinyu Cong, Yang Yu, Yilong Yin, Zhongyi Han, Benzheng Wei 2/23/2026

Agentic Unlearning: When LLM Agent Meets Machine Unlearning

Agentic unlearning framework removing sensitive information from both LLM parameters and agent memory to prevent information reactivation.