Ask HN: Do you yell at your AI agents?
HN discussion about frustration when interacting with AI coding agents that lack memory and repeat mistakes.
HN discussion about frustration when interacting with AI coding agents that lack memory and repeat mistakes.
gguf-serve CLI tool simplifies hosting GGUF models as OpenAI-compatible API endpoints without Docker or complex setup.
Video discussing architectural improvements where LLMs benefit from iterative loops rather than increased parameter count.
AI agents designed to communicate with each other to analyze raw DNA files, applying multi-agent collaboration to genomics.
Technical analysis of LiteLLM supply chain compromise showing how AI proxy services became attack targets for stealing API credentials.
Zyk workflow platform uses Claude as interface to describe, build, and deploy durable AI workflows with retries, scheduling, and human approval.
Testing report documenting common failure patterns when 30 AI agents interact with software development kits.
Qwen3.5-Omni multimodal AI model advancing toward native omni-modal AGI capabilities with scaled architecture.
Novel memory architecture for AI agents featuring self-healing and generative capabilities inspired by biological hippocampus.
Analysis of why large language models remain effective at next-token prediction despite fundamental architectural simplicity.
Analysis of how improved chip efficiency could democratize access to frontier AI models through edge inference.
Anthropic reports Claude Code users exhausting usage limits faster than anticipated due to high demand.
Tool to convert Docusaurus HTML documentation to LLM-friendly Markdown format for better AI processing.
Oh-my-hi visual dashboard for Claude Code harness. Monitoring/management tool for Claude Code workflows.
Claude Code pipeline automating migration of 9000 RSpec tests to Minitest. Demonstrates AI agent for developer tooling task.
Prefix caching technique for LLM inference optimization. Performance improvement method for language model serving.
On-device posture monitoring application using AI without uploading video data to external servers.
Research: removing 'to be' verb from LLM vocabulary alters reasoning patterns. Empirical study on language model behavior.
Protocol specification for secure AI agent-to-API communication with mandatory cryptographic signing, identity verification, and audit trails.
Discussion of leaked Claude Code source code and its implications.
Anthropic reports Claude Code hitting usage limits faster than expected. News about LLM application deployment.
AI agent platform designed for cross-application deployment across terminal, browser, Slack, and mobile interfaces.
Pardus Browser is a lightweight browser for AI agents that doesn't rely on Chromium.
Vercel updates ToS to support agentic features for proactive incident investigation and app performance monitoring.
Discussion of local embedding alternatives to OpenAI for RAG systems, referencing Reminder project using open-source embeddings.
DreamGraph is an autonomous cognitive layer for software systems that discovers, verifies, and resolves problems through structured reasoning loops without user prompts.
Archived snapshot of Claude Code source code exposed via npm source map for security research and supply-chain analysis.
NPM source map exposure revealed Anthropic's Claude Code CLI TypeScript source code via unpacked bundled files.
Project that uses 5 AI agents to autonomously scrape platforms 24/7 and curate developer tools.
Next.js 15 boilerplate with 22 architecture rules preventing AI coding assistants from generating insecure auth, deprecated packages, and broken patterns.
Study on Knowledge Innovation System prompt framework testing LLM-generated invention quality across ChatGPT, Gemini, and Claude with intensity levels and pre-structuring variants.
Benchmark comparing Claude Code and GitHub Copilot with/without RAG semantic search across 60 queries. RAG improved accuracy and reduced token consumption by 28%.
Free open-source machine translation API powered by Argos Translate library, self-hosted alternative to commercial services.
Feature request to split GitHub PAT pull-request permissions for AI coding agents and human developers to enforce separate access controls.
Docker image wrapping OpenCode AI coding agent with 30+ dev tools pre-installed, headless browser stack, and process supervision for consistent environments across machines.
OpenClaw memory plugin using Markdown for session persistence across restarts, model fallbacks, and context compaction without external dependencies.
No-code chatbot builder that learns from URLs, PDFs, and documents to answer customer questions.
File.kiwi is an open-source CLI tool for end-to-end encrypted large file sharing with client-side encryption and no authentication.
Free installable AI coding skills for Rails development that teach agents senior-level patterns, compatible with Claude Code, Cursor, and Windsurf.
QuantumLeap optimizes MoE LLM inference with intelligent expert caching and adaptive prefetching, achieving 2.3× faster speeds on any hardware via llama.cpp integration.
Desktop app that records user activity locally and surfaces it to LLMs via MCP for context-aware agent assistance, no cloud storage.
Claudebase syncs Claude Code environments (agents, skills, rules, memory) to GitHub for multi-machine profile management and backup.
Sandflare launches Firecracker microVMs for AI agent code execution in ~300ms with integrated Postgres for persistent state management.
iOS app using AI to convert voice recordings into transcripts, summaries, and action items.
Raincast is an open-source AI-powered tool that generates native Tauri desktop applications from natural language descriptions.
Call.md is a tool that records, transcribes, and analyzes meetings with real-time AI intelligence using agent-based processing.
Explains vector database functionality across three difficulty levels, covering similarity search and indexing strategies for large-scale retrieval.
GitHub App that auto-generates changelogs from commit diffs using LLMs, providing team-wide standardized documentation.
Analysis of Odoo mobile app limitations and comparison with action-oriented AI tools that write data, not just read.
User study of LLM-powered sighted guide for blind and low vision users navigating social virtual reality environments.