Show HN: Kronaxis Router – Don't pay frontier prices when a local LLM is enough
LLM router that prioritizes local models over frontier APIs to reduce costs. Title only, assumes cost-aware LLM routing.
LLM router that prioritizes local models over frontier APIs to reduce costs. Title only, assumes cost-aware LLM routing.
Desktop dev workspace integrating Claude with Kanban board, multi-repo support, and agent SDK for improved task management and iteration.
Chrome DevTools MCP enables AI systems to interact with browser debugging tools for enhanced perception capabilities.
Vix: AI coding agent achieving 50% cost reduction and 40% speedup vs Claude Code using virtual filesystem minification and stem agents for cache optimization.
Spacebot is an agentic AI system with a dedicated LLM process role, positioned as OpenClaw alternative.
Hotcopy is a CLI tool with AI agents that maintain context across sessions and learn over time; developer tool for coding.
Mobile IDE for SSH server management and AI agent orchestration across servers from iPhone interface.
Computer vision ML techniques for room occupancy detection; machine learning application with limited scope.
Self-building knowledge graph of contradictions using local LLM for analysis and visualization.
Project Glasswing uses Anthropic's Claude Mythos 2 frontier model to identify software vulnerabilities, demonstrating AI capabilities in cybersecurity.
Discussion of security requirements as AI models become more capable, using Anthropic's Mythos as case study.
Self-hosted AI research agent that browses web, takes notes, and writes reports with crash recovery via Postgres checkpointing.
Opinion on low-quality AI-generated content; lacks substantive technical analysis.
Project Glasswing uses Anthropic's Claude Mythos model to identify software vulnerabilities; demonstrates AI coding capability for security applications.
Development environment and runner for managing prompts, workflows, and agent pipelines at scale with dataset testing.
AI-powered product feedback tool analyzing submissions and providing detailed improvement suggestions.
Open-source toolkit enabling AI agents to collaborate with humans in reactive marimo Python notebooks with working memory.
Discussion of GitHub Copilot vs alternatives like Cursor and Claude Code for code completion and agentic features.
Rust async client library for Ollama local LLM API with streaming, chat, and embeddings support.
CLI + MCP server tool enabling coding agents visual verification of UI layouts via browser-based testing.
Technical guide on prompt caching optimization techniques with AI co-authoring.
DuckDB-based database system for SQL-capable AI agents with benchmarks across 11 LLMs.
Discussion on marketing developer tools built with rapid development practices.
Tool that automatically discovers optimal system prompts for LLM tasks by analyzing desired output examples, eliminating manual prompt engineering.
OS-level containment system for AI coding agents on macOS, addressing security risks when running untrusted agent code with filesystem/system access.
Developer tool that creates queryable knowledge bases from videos/podcasts for AI agents.
Podcast discussion on governed AI systems in healthcare domain.
Discussion of security vulnerabilities in AI agent sandbox implementations.
macOS containment system using kernel sandboxing and firewalls to safely run unrestricted Claude Code agents autonomously.
AI assistant prototype with session-aware memory that forgets context when users leave.
Best practices guide for AI agent guardrails, covering pre/post-LLM safety patterns.
Mendral is a CI specialist coding agent built on Claude. Demonstrates how identical LLMs produce different outputs through system prompts, tools, and context optimization.
Critique of Anthropic's own AI implementation practices versus enterprise recommendations.
Compares local LLMs (Gemma4-26B, Qwen3.5-35B) for agentic coding tasks using OpenCode and Pi-Coding-Agent with custom tool usage scenarios.
A wiki-based LLM system built on CIS security controls documentation, enabling semantic search and knowledge retrieval over structured security frameworks.
Open source modular OS framework for designing and deploying AI agents. Show HN submission with practical agent infrastructure.
ErrataBench is a benchmark measuring LLM proofreading performance across 51 model variants using an agent loop, with detailed runtime and cost metrics.
Discussion of LLM collaboration patterns in developer tools like Cursor and Claude. Explores user experience challenges with autonomous AI agents.
Infrastructure platform for payment processing integrated with AI agents in EU.
Testreel: npm package for programmatic demo video generation from JSON/YAML/Playwright. Enables LLM agents to create product demos with cursor overlay and customizable backgrounds.
Tutorial on building AI agent for Slack using Chat SDK and AI SDK. Developer guide for LLM integration.
Multi-agent system organized as functional company with independent AI agents in HR, engineering, design roles. Novel agent architecture approach.
Open source framework extracted from 500+ production AI agents. Production-tested patterns and tools for building agents at scale.
Case study: Using AI to reconstruct and resurrect a 1992 MUD game from artifacts and old documentation. Technical restoration project.
SQLite extension providing persistent searchable memory for AI agents with vector search, markdown support, and offline-first sync. Open source.
Open-source CLI coding agent supporting multiple LLM providers (OpenAI, Gemini, Ollama, local models) with MCP and streaming.
CLI tool for querying JSON/JSONL files, specifically built for AI agent workflows. Open source developer utility.
Lemonade 10.1 release with optimizations for running local LLMs on AMD GPUs and NPUs.
Research comparing LLM performance differences between API-driven and GUI-based (touchscreen) interaction modes.
Discusses decentralized training approaches to reduce AI model training energy consumption.