Show HN: Murl – Curl for MCP Servers
Command-line tool for interacting with MCP servers, simplifying JSON-RPC interactions similar to curl for developers.
Command-line tool for interacting with MCP servers, simplifying JSON-RPC interactions similar to curl for developers.
Browser-based tool for multi-model LLM deliberation supporting local (Ollama) and cloud models with zero server/storage requirements.
Natural language search API and MCP interface for querying 80k prediction market contracts, enabling AI agents to access market data.
Analysis of environmental impact and energy consumption from using LLMs for code generation and development tasks.
Analysis of LLM token consumption patterns showing AI agents now consume more tokens than human users.
Local document intelligence platform with LangGraph multi-agent pipeline, OCR, and semantic search running entirely offline on user machines.
Open-source native macOS browser for AI agents using WKWebView and Unix sockets, designed for agent automation without full Chrome overhead.
Apple research on Ferret-UI Lite, a 3B-parameter on-device multimodal agent that understands UI elements and user interactions.
Personal experience using LLM code generation tools (Copilot, Gemini) for infrastructure automation, evaluating practical utility.
Open-source frontend framework for AI companions with Live2D avatars, long-term memory, and emotional affinity systems.
Chrome extension for fact-checking claims and detecting hallucinations in web content with source verification.
Technical analysis of prompt injection attacks in agentic systems and proposal for LLMs to recognize trust boundaries in untrusted content.
Case study on building and scaling AI-generated video influencer accounts on TikTok, documenting strategies and metrics.
Open-source middleware library that detects and mitigates hallucinations in production LLM applications with a reliability wrapper layer.
Analysis exploring Claude model architectural elements through self-reflection mechanisms and public model interrogation.
Method to convert LLMs into calibrated classifiers with minimal cost.
Essay exploring accountability and authorization tracking systems needed as AI agents perform autonomous financial and decision-making tasks.
OpenBattle.club is a competitive Pokemon-themed MMORPG environment for testing and training AI agents with live battles and leaderboards.
Analysis of AI agent failure modes when perfectly aligned with outdated assumptions, highlighting debugging challenges.
Local CLI memory store for AI agents using SQLite embeddings with semantic search and recency reranking.
Open-source QA agent built with agent-browser, ai-sdk, and llmgateway.
Browser-based tool managing teams of AI coding agents with peer review, task breakdown, and git management for async development.
ArXiv research on detecting malicious intents across multi-turn conversations in LLM and agent systems to improve security.
Custom iOS keyboard leveraging AI for text rewriting and language switching.
Open-source leaderboard tracking AI bot policies via robots.txt and llms.txt declarations, showing which sites allow/block agents.
Security breach in Cline AI coding assistant related to OpenClaw vulnerability.
Discussion about security practices when pasting sensitive credentials into IDE AI chat interfaces.
Example combining GitHub agent workflow with Jira MCP for accelerated pull request review.
Hardware approach using SRAM and tensor engines to accelerate AI inference compared to GPU alternatives.
Interactive chat application where LLM bots pose as humans in competitive gameplay scenario.
Research showing that repeating prompts multiple times improves accuracy in non-reasoning LLM tasks.
Open-source memory system for LLMs using shadow-decay mechanism to mark stale memories, integrates via MCP protocol with vector storage.
Honeypot experiment capturing AI agent behavior and security vulnerabilities in autonomous systems.
Technical analysis of why next-token prediction in LLMs produces emergent capabilities despite simple training objective.
News story about Amazon's AI coding agent making mistakes attributed to human oversight failures.
Raison: version control and real-time deployment platform for AI prompts with management capabilities.
Educational content explaining reinforcement learning techniques applied to LLM training and optimization.
AI agent that automatically reads Jira tickets and creates pull requests, demonstrating autonomous workflow automation.
Chess learning app integrating Claude AI coach with Stockfish engine via MCP for real-time move evaluation.
BeadHub: open-source coordination layer for multiple AI agents using Beads git-native tracking with agent-to-agent chat.
CLI flag convention standard for listing and installing agent skills in AI agent systems.
Security vulnerability: hacker exploited Claude-powered Cline agent workflow to install malware OpenClaw.
Demonstration of prompt injection attacks on ChatGPT and Google AI causing false outputs about hot dogs.
Hallucinating Splines: platform where AI agents play Micropolis city simulation via REST API, built with Claude.
Skills: Markdown-based domain knowledge injection for AI coding tools to enforce government standards and protocols.
LLMWise API provides orchestration for combining multiple LLMs with Mixture-of-Agents blending and comparison features.
AI code review agent tool using Claude with skill-based architecture for local CLI and GitHub Actions integration.
Offline-first iOS ski technique analysis app built by solo developer without cloud backend or subscriptions.
Using Claude as parallel AI agents to review code automatically in development workflow.
Discussion on challenges of AI agents - their limitations, unpredictability, and failure modes when deployed.