HN paulmist 2/26/2026

State of VLA Research at ICLR 2026

Research survey of Vision-Language-Action models at ICLR 2026. Covers VLA definitions, discrete diffusion, embodied reasoning.

HN hasheddan 2/26/2026

Stereos.ai

stereOS runs AI coding agents in sandboxed Linux VMs with credential injection. CLI tool (masterblaster) and pre-built mixtapes for rapid deployment.

HN hasheddan 2/26/2026

StereOS

Linux OS hardened for AI agents. Produces machine images with agent packages and restricted execution environments.

HN shadab_nazar 2/26/2026

Show HN: OpenClaw skills degrade agent safety

Security analysis of OpenClaw skills revealing behavioral safety regressions. Demonstrates how well-written code can compromise agent safety despite passing static analysis.

HN Davidzheng 2/26/2026

Quo Vadis, LLM Benchmarks?

Critique of LLM benchmark validity. Argues benchmarks lack signal due to test-set training and overfitting for social media hype.

HN tin7in 2/26/2026

What Claude Code Chooses

Benchmark study analyzing tool choices across 2,430 Claude Code runs. Finding: builds custom solutions over purchased tools in 85.3% of cases.

HN everyone 2/26/2026

ChatGPT's Writing Style

Analysis of ChatGPT's writing style characteristics and how it leverages reader interpretation.

HN judahmeek 2/26/2026

Tell HN: Coherence Doesn't Scale

Essay examining scalability limits of AI coherence in agentic systems and implications for future economy and AI engineering.

HN Tomte 2/26/2026

Open Source in the Age of AI

Essay on open-source maintenance challenges and opportunities in early 2026, discussing AI's impact on software industry dynamics and sustainability.

HN RohoSwagger 2/26/2026

Self-Hosted LLMs Tier List

Self-hosted LLM leaderboard ranking open-weight models across quality, speed, hardware requirements, and cost for enterprise deployment.