Show HN: Arrivl – Analytics for AI agent traffic on your site
Show HN: Analytics tool detecting AI agent traffic from server logs, addressing JS-blind AI traffic in analytics platforms.
Show HN: Analytics tool detecting AI agent traffic from server logs, addressing JS-blind AI traffic in analytics platforms.
Show HN: Game development project using Claude to build complete games in single prompts. Chess game demonstrates working AI opponent.
Technical overview of serverless GPU infrastructure for running inference on large language and neural network models at scale.
Building AI agents capable of local Android emulation for automated testing/interaction.
Performance optimization system for AI agent harnesses using Claude Code.
RAR archive implementation in Rust using LLMs (Codex, Claude Opus). 55k lines completed in 5 weeks, demonstrating LLM-assisted reverse engineering.
Using LLMs to detect and identify bugs in Python C-extension code.
Event-sourced logistics integration bridge that normalizes fleet telemetry and exposes data via MCP tools for LLM agents.
Multi-agent system for data engineering, analysis and statistical reasoning that adapts to existing data stacks or builds new ones.
Anthropic increases Claude Code weekly usage limits by 50% through July 13.
Open-source tool to find warm introductions from contact networks by matching LinkedIn profiles, alternative to paid services.
Red Hat's skill packs enable AI agents with institutional memory access. Agent tooling feature unclear.
Research comparing Exa vs Google as search backends for reinforcement learning trained agents, showing RL performance improvements.
Tutorial series building AI agents from scratch to explain how agents work, covering Claude, Codex, and related tools.
News that Apple is developing policies to allow AI agent applications on App Store.
Research on predicting rare LLM failures using 30× fewer rollouts for testing efficiency.
Security researcher found 500 vulnerabilities in apps using LLMs for detection. Title only, minimal details.
MCP server for Claude/ChatGPT providing access to lens/CMOS camera dataset for optics selection in AI applications.
Technical explanation of how language models work mechanically, covering transformer internals and token prediction.
Announcement of K2 image generation model optimized for aesthetic output.
Benchmark for evaluating customer-facing AI agents on real-world knowledge base navigation and application tasks.
VibeLoom tool for AI-assisted code generation using contract-driven development.
Research on LLMs that combine fast and slow learning mechanisms for continual adaptation.
Profine tool for automated profiling and code rewriting optimization in ML training loops.
Engineering case study of AI coding agent that opens production pull requests for payment integrations, addressing undocumented quirks.
AgentDeck: game console environment for AI agent research and testing.
Ardent (YC) builds database sandboxes for coding agents to test safely without production risk.
Guide to reviewing agent-generated pull requests, technical debt patterns, and code quality issues in AI-generated code.
Archivists using LLMs to decipher handwritten historical documents at scale for digital preservation.
User review of Claude Code 2.1.139's new Agent View and background sessions, noting useful features with remaining rough edges.
Audrey - local-first memory management system for AI agents with source code available.
Discussion of complexity in agentic systems beyond prompt optimization, noting multiple moving parts.
EleutherAI's LM Evaluation Harness - benchmarking framework for evaluating language model performance across tasks.
BossHogg is a PostHog CLI for AI agents enabling HogQL queries, feature flags, and insights from terminal or Claude Code.
Ledger is a local Rust tool analyzing Claude Code token spend via JSONL logs with CLI and dashboard.
Opinion piece on competitive challenges of building products with AI coding assistants becoming accessible to all.
Ratify Protocol - cryptographic method to prove AI agent authorization offline with sub-millisecond verification.
Business ideas for AI infrastructure including service connectors, CDN, and database integrations for AI assistants.
Technical overview of Anthropic's Claude Managed Agents released April 2026: hosted agent infrastructure with sandboxed environments.
Kubernetes cost analyzer using Claude API to identify waste by namespace/pod without agents or dashboards.
Microsoft MDASH - multi-model agentic security system that discovered 16 vulnerabilities including 4 critical RCE flaws in Windows.
Endpoint Context Protocol converts HTML to Markdown for AI agents. Developer tool for agent web interaction.
ADK framework for building persistent AI agents with pause/resume and context preservation. Developer tool for agent development.
AgentGate authorization framework for AI agents. Developer tool for access control in agent systems.
Scherlok data quality monitoring tool using ML to detect production data anomalies, integrates with dbt.
Critical analysis of OpenAI's creative writing LLM, examining stylistic patterns and limitations in generated text.
Proposed governance standard for autonomous AI systems deployment.
Tool for generating 3D objects with separate, editable parts using LLMs (Gemini, Claude, ChatGPT) to create Blender construction scripts.
Research code and findings from cracking Jane Street LLM models with mechanistic analysis and ablation studies.
Analysis of production system failures when AI-generated code reaches deployment, from applied AI experts.