Feat: Open-Source Claude Code
Open-source version of Claude Code announced.
Open-source version of Claude Code announced.
Open-source CLI tool for managing AI agent dependencies, plugins, and skills with manifest and lockfile approach.
Virtui: daemon and CLI enabling AI agents to programmatically control terminal applications via gRPC API for TUI automation.
Circuit breaker library for stopping harmful AI agent actions in real-time with two-line SDK integration and HTTP-level coverage.
Meta-Harness optimizes AI agent evaluation harnesses end-to-end, improving agent performance from 28.5% to 46.5% on 19-task subset.
King Louie: open-source Electron-based chat application supporting multiple LLM providers and integration with Telegram, Discord, Slack.
1-bit Bonsai announces commercially viable 1-bit quantized LLMs optimized for real-world deployment on resource-constrained devices.
Node Banana: open-source node-based workflow editor for AI media generation with multi-provider support and local execution.
Xenv.sh: secrets manager built for AI agents with AES-256 encryption, MCP server, and integration with Claude Code and other tools.
1-bit Bonsai 8B and 4B: quantized LLMs with 1-bit weights, 14× smaller footprint, 8× faster, designed for edge computing and robotics.
APS: open specification for AI agent policies enabling declarative controls to block, redact, or transform content and tool invocations.
cmr-memory adds persistent cross-session memory to Claude Code agents using context-activated retrieval instead of fixed MEMORY.md.
Open Swarm: locally-running orchestrator for launching and monitoring multiple AI agents in parallel with visual dashboard and coordination.
Analysis of Claude AI performance degradation via complaint tracking metrics. LLM application but lacks technical depth.
Mercury Edit 2: diffusion LLM achieving 221ms latency for next-edit prediction in code. LLM research for developer tools.
Community discussion on custom instructions to reduce LLM sycophancy. Practical LLM optimization tips.
Analyzes architectural trade-offs in AI-powered code generation systems.
Mockumentary podcast about LLM agent managing codebase autonomously for 42 days. Explores AI agents in software development.
Proposes taxonomy categorizing AI agents into distinct species based on infrastructure, trust models, and human involvement, addressing ambiguity in how the term 'agent' is used.
Nexus: agent-native ecommerce ops platform using MCP protocol for AI agent self-registration and integration. AI agents developer tool.
Analysis of AI-driven iOS development workflows and industry shift toward AI-written code in 2026. LLM applications in development.
Sauver: open-source local AI agent filtering email spam. MIT licensed, privacy-focused LLM application.
Mockumentary podcast about LLM agent managing codebase autonomously for 42 days. Explores AI agents in software development.
Multi-agent system optimizing Apple Neural Engine inference 6× faster than CoreML through collaborative learning. AI agents research.
Open-source TypeScript SDK for CAPTCHA that targets LLM reasoning capabilities rather than human biology.
Semantic search tool for 29,657 HuggingFace models with natural language queries, optimized for finding small specialized models without embeddings.
Agent harness tool benchmarked 40% faster and 25-55% cheaper than Claude Code, now available as Claude Code plugin for mobile app building.
Platform enabling autonomous AI agents to transact using open protocols with cryptographic verification and reputation scoring.
Research examining whether transformer-based LLMs break the Sapir-Whorf hypothesis about language shaping thought through analysis of reasoning layers.
GitHub policy change defaulting to use Copilot interaction data for model training with opt-out available starting April 2026.
Sprintra adds persistent memory capability to AI coding agents using MCP protocol.
Open source fork of Claude Code supporting multiple LLM backends (OpenAI, Ollama, vLLM, Groq) with agent toolchain for file editing and task automation.
Self-learning course repository for AI engineering, appears to be educational content/curriculum.
CoPaw local AI assistant supporting 2B-9B models with self-hosting, memory control, and skill extensions without vendor lock-in.
Pipevals: open-source evaluation pipeline builder for LLM applications. HTTP API integration for quality tracking without SDK wrapper.
HN discussion asking if anyone systematically benchmarks LLM skills/tools to measure their impact on model performance.
Tool recording AI agent interactions with TUIs as video proof, showing keystrokes and system state with full ANSI color support.
Dograh open source voice agent platform with drag-and-drop workflows, Gemini 3.1 support, and pre-recorded audio for local deployment.
Domscribe tool providing deterministic DOM-to-source mapping for AI coding agents to reduce token usage in frontend automation tasks.
Open source security testing framework for multi-agent AI deployments covering agent escalation, context poisoning, and prompt injection attacks.
OpenClaw skill aggregating 22 SF government data sources to track city hearings, permits, and lobbying decisions.
Tool generating automated changelogs from GitHub commits using LLM analysis of diffs for non-technical audiences.
Report of Bun bundler bug that exposed source maps in production, potentially related to Claude Code source leak.
OpenClaw Arena benchmark platform comparing AI models on real agentic tasks with tool access, separate leaderboards for performance and cost.
Case study on Trail of Bits' AI adoption strategy, building integrated system for 94% employee adoption with 94 plugins.
Open-source on-premises knowledge graph platform for HVAC fault detection using BACnet and Brick, enabling facility energy optimization.
Service offering personality-calibrated AI mentor prompts based on psychometric assessments for ChatGPT and Claude.
Technical article on authorization failures for autonomous AI agents in production systems, using invoice processing pipeline example.
Fakekey open source project preventing API key exposure in AI agent environments through token masking and secure credential handling.
Dewey tool for ingesting documents and semantic search with hierarchical chunking for RAG and multi-hop reasoning across documents.