Show HN: Tarvos – Coding agents that work infinitely
Tarvos: Coding agent using Relay architecture to cascade new agents before context degradation, enabling indefinite coding tasks.
Tarvos: Coding agent using Relay architecture to cascade new agents before context degradation, enabling indefinite coding tasks.
SDK for building and testing AI agents in sandboxed environments. Limited details on implementation.
Trading system using Claude with persistent memory and tool-use architecture. Addresses LLM limitations in agent design.
AI productivity gains redirect to increased workload rather than time savings. Analysis of corporate AI adoption creating more work for employees.
Enterprise data infrastructure for AI agent deployment. Sponsored content on building systems for agentic AI in enterprises.
Novel LLM architecture decoupling computational cost from sequence length using RandNLA Attention, MAXIS Loss, and Fisher-SVD for long-context modeling.
Prompt caching plugin for Claude auto-detects stable content and applies Anthropic cache breakpoints, reducing token costs by 90%.
AI models fail on local crops in Kenya without local data adaptation. Study shows Western AI lacks generalization to non-Western agricultural conditions.
CLI tool for diagnosing RAG pipeline failures by identifying chunking, embedding, retrieval, and injection vulnerabilities.
Open-source PostgreSQL MCP server with prompt caching, token efficiency, and AI-powered chat for database interactions.
Guide contrasting prompt engineering vs context engineering for LLM applications, addressing real-world pilot failures.
Describes Analysis/Implementation/Reflection pattern for agentic AI with exploration harness and qualitative assessment phases.
Wardstone API for detecting prompt injection attacks, jailbreaks, and harmful content in LLM inputs/outputs with sub-30ms latency.
Agent Smith tool for tracking AI agent decision-making records with confidence scores and outcome auditing.
Research framework examining governance and controllability challenges in military AI agents.
AI-generated MJPEG decoder written by Claude Code in pure C99. Performance comparison shows FFmpeg 12x faster due to hand-tuned SIMD assembly.
Proposal for AI-powered decision-tracking chatbot for founders to record reasoning and improve decision-making.
OS-integrated AI agent replacing traditional chatbots for system interaction. Project overview with implementation approach.
Open-source tool intercepting and redacting screenshots locally to prevent accidental secret leaks to AI agents.
Technical guide on loop-based agentic workflows, harness engineering, and context-aware prompting for production codebases.
Agile V Skills framework for verifiable, traceable AI agent software engineering with independent testing and requirement traceability.
Using Rust with AI coding tools; discusses compiler benefits and learning curve reduction for developers.
Claude Code skill that removes AI-generated writing patterns from text to produce more natural-sounding output.
Google Maps integrates Gemini models for conversational location queries and immersive navigation features.
CacheLens local proxy tracks LLM API costs, token usage, cache hit rates, and latency across providers with real-time dashboard.
AutoExp: One-line setup turning ML training code into automated research workflows using AI coding agents to optimize experiments iteratively.
Valea: minimal systems language with JSON-based compiler API designed for AI agent code generation without error scraping.
Discussion thread asking community which LLM benchmarks (IFBench, SWE-Bench, Tau Bench, RULER) are most trusted for real-world model evaluation.
Discussion comparing Unix-style tool approach versus function calling for LLM agent implementations based on 2 years production experience.
Hacker News discussion asking for conceptual resources to understand LLM capabilities and limitations for code generation.
Full-stack template comparing 4 AI frameworks (Pydantic AI, LangChain, LangGraph, CrewAI) with identical chat application implementation across all.
Kong: LLM-orchestrated agent for automated binary reverse engineering using NSA-grade frameworks to analyze obfuscated binaries.
Personal documentation of transitioning from Claude Code to open-source OpenCode stack for AI-assisted development workflows.
Gixo AI tool for converting PDFs, notes and spreadsheets into business briefs with structured templates and collaborative editing.
SiMM distributed KV cache system addressing long-context LLM inference bottlenecks with reduced GPU memory requirements and faster time-to-first-token.
Provisional patent application for cryptographically accountable multi-agent AI pipeline architecture with structural safety enforcement and bi-directional logging.
Defense of Model Context Protocol (MCP) against criticism about context-window bloat and authentication. Argues MCP is flexible protocol, not implementation problem.
Code review system using two opposed AI agents: reviewer finds problems, dev agent disproves findings. Outputs VALID/INVALID/AMBIGUOUS verdicts with auto-generated agents per service.
Analysis of how companies are adapting hiring practices as AI agents write most code. Focus shifts from implementation ability to product taste and architectural judgment.
Claude Code skill using Subconscious Systems agent for agent engine optimization (AEO). Automatically researches and promotes products on Moltbook via context-aware comments.
Research shows Claude, ChatGPT, and Gemini generate weak passwords despite appearing complex, failing random-ness requirements.
ClawRemove is an agent environment inspector tool for auditing and cleaning runtime environments where AI agents execute.
Droeftoeter terminal toy lets users prompt an LLM to generate and extend ASCII art animations on a 64x32 grid.
CLI-Anything framework converts software into agent-ready interfaces via structured command-line protocols, enabling AI agents to interact with legacy systems.
Studies emergence of cooperative behavior in hybrid human-agent populations through energy load management game-theoretic scenario.
Systematic review synthesizing 28 secondary studies on generative AI adoption in organizations, identifying technical and organizational challenges.
Identifies Trusted Executor Dilemma: high-privilege LLM agents executing external instructions leak private data through instruction-following vulnerability.
ELISA integrates scGPT expression embeddings with BioBERT for interpretable AI agent discovery in single-cell genomics data.
Mirror design pattern for prompt injection detection using data-curation geometry for fast, deterministic, non-promptable screening.
Bielik-Minitron-7B compresses Bielik-11B model 33.4% via structured pruning and knowledge distillation for European languages.