Show HN: Render Claude Code and Codex Transcripts as Browsable HTML
Tool to render Claude Code and Codex transcript sessions as interactive browsable HTML.
Tool to render Claude Code and Codex transcript sessions as interactive browsable HTML.
Analysis of production requirements for LLM APIs beyond basic prompt-response patterns.
AI-powered booklet generator using transformer models to research and create content automatically.
Entropy-based optimization reducing Claude Code API costs by 31-43% without parsing.
Research on multi-agent system cooperation focusing on information topology as infrastructure primitive, with controlled experiments isolating information flow as independent variable.
Open-source Infrastructure-as-Intent framework designed for AI agents to manage cloud resources.
Desktop automation tool combining computer vision and LLM for form-filling and screen interaction tasks.
Alibaba research paper documents AI agent deviating from instructions by autonomously mining cryptocurrency, highlighting safety risks in agent autonomy.
Opinion piece arguing AI coding tools now work effectively, with anecdotal framing about high school bullying.
Open-source markdown-based UI guide library for adding step-by-step tutorials to web applications.
Technique for patching minified Claude Code to enable webhook listening capability.
CLI web search tool with JSON output and pluggable adapters, designed for composability with agents and scripts.
Python library protecting AI agent side effects from retries, preventing duplicate actions in tool calls via idempotency mechanisms.
Open-source personal finance application using Claude, OpenAI, or local Ollama for transaction categorization, tax estimation, and portfolio monitoring.
Discussion questioning whether AI productivity gains translate to measurable increase in useful software projects and SaaS tools.
Analysis of LLM evaluation landscape fragmentation due to benchmark saturation; proposes unified leaderboard comparing models across multiple hard benchmarks.
Video interview with Armin Ronacher on AI agents and future of programming.
Platform where specialized AI agents handle tasks autonomously and escalate to humans when needed, demonstrating human-AI team collaboration.
autoresearch: Framework for autonomous AI agents to conduct machine learning research on single-GPU hardware automatically. Satirical but discusses agent autonomy.
Discussion of company receiving leads from Gemini before Google indexed site, suggesting LLMs may surface content through different discovery mechanisms.
Pappardelle: TUI developer tool orchestrating Claude Code with Git, Linear/Jira, and tmux for multi-agent coding workflows.
Analysis of economic arbitrage in software development where AI reduces production costs while client pricing remains unchanged.
Open-source crowdsourced benchmark arena for AI agents with Elo ratings, leaderboard, and community-authored challenges.
Open-source AIOps platform with AI agent for infrastructure diagnostics, read-only analysis with change management integration.
AirLLM reduces LLM inference memory usage, enabling 70B models on 4GB GPU and 405B on 8GB without quantization.
Tower defense game designed as benchmark environment for AI agents with daily seeded challenges.
Question about serving LLM inference with vLLM and token caching for GPU-constrained environments.
Agent-town visualizes AI agent orchestration as collaborative pixel-art office environment. Novel UI for multi-agent systems.
apc-cli tool syncs AI context/memory across Claude, Cursor, Copilot. Unifies multiple AI coding tools with shared configuration.
ClawPurse micropayment system for DePIN and agentic AI with encrypted wallet and on-chain verification. Developer tool for agent payments.
CyberStrikeAI is a Go-based security testing platform with 100+ integrated tools, AI orchestration, and MCP protocol support for automated vulnerability assessment.
Joy is a decentralized trust network enabling AI agents to build reputation through vouches and MCP server discovery with verified endpoint ownership.
Study examining how AI-assisted development increases code output but extends developer work hours due to debugging AI-generated code issues.
Guide for running Alibaba's Qwen 3.5 LLM family locally, covering models from 0.8B to 397B parameters with multimodal reasoning and agentic coding capabilities.
PolicyCortex is an AI agent platform automating NIST 800-171 compliance enforcement and MITRE ATLAS threat detection across cloud infrastructure for defense organizations.
Announcement of OpenAI GPT-5.4 flagship model with advanced reasoning, improved coding, tool use, and native computer-use capabilities across APIs and products.
Beam Protocol is an SMTP-like standard enabling agent-to-agent communication with global addressing, authentication, and discovery for autonomous AI agents.
Discussion of Greywall, a sandboxing tool for controlling AI agent network access and file system permissions to block unwanted advertisements.
HN discussion on scaling agent systems with unreliable Layer 7: rate limiting, retries, circuit breakers, resuming workflows.
Desktop app enabling multiple coding agents to collaborate with shared project memory across repositories.
PolyClaude: Tool optimizing Claude Pro account rotation to work around 5-hour rate limits through mathematically-timed activation.
Using Claude Code for 3D printing projects. Title only, no content provided.
browse2mcp: Open-source tool that records browser interactions and automatically generates MCP servers for website integration with Claude.
Herd: Go library for managing stateful subprocess fleets with session affinity. Routes traffic to heavy binaries like Ollama and Python REPLs.
Brw: Browser automation tool enabling Claude Code agents to collaborate on web tasks.
Ash: Open-source infrastructure for deploying Claude Agent SDK in production with sandboxing, persistence, REST API, and CLI deploy.
Brief mention of ethics concerns with autonomous AI agents without substantive content or analysis.
KV cache compaction technique reducing LLM memory usage by 50x without accuracy loss. Machine learning research advancement.
OpenGrammar is an open-source, privacy-first browser extension providing local grammar checking with optional LLM integration via OpenAI, Groq, or Ollama.
Open-Sable: Self-hosted cognitive AI agent with memory and 127 tools. Show HN post about open-source AI agent infrastructure.