Ask HN: what’s your favorite line in your Claude/agents.md files?
HN discussion sharing prompt engineering tips for Claude with emphasis on detailed planning and testing strategies for LLM-assisted development.
HN discussion sharing prompt engineering tips for Claude with emphasis on detailed planning and testing strategies for LLM-assisted development.
Practical guide on using GPT-5.4 for generating production-ready frontend designs with improved UI capabilities and image understanding.
Title only. Likely about robust AI agents. No abstract or details available.
Essay on programmer psychology regarding AI tool adoption. Argues process vs. result matters for adoption.
Analysis of Stripe's 1,300 PR/week AI claim reveals 0.37 PRs per engineer weekly, questioning actual productivity gains and methodology.
Page-Agent.js open source tool enables natural language control of web interfaces using LLMs. Includes demo and documentation.
Essay on AI exceeding human cognitive capability and implications for cognitive labor displacement. Personal reflection from Wispr CTO.
Sandbox workspace tool designed for AI agents with terminal interface, supporting Claude and other LLM runtimes.
AI-written article on history of computerized education proposals. Discusses dystopian AI replacement narratives in education over 50+ years.
LessTokens offers AI-powered prompt compression to reduce LLM token usage by 25% without quality loss. Simple API integration.
Observation: ChatGPT exhibits bias toward selecting numbers in 7200-7500 range when asked to pick 1-10000, suggesting non-uniform sampling.
Elastik: Framework treating LLM as untrusted HTTP client using MCP protocol. Enables LLMs to build web apps safely in <200 lines of code.
S2S validates motion training data using physics-based principles before ML training. Detects corrupted/synthetic data violating Newton's laws and kinematics.
Double-o: Context-efficient command runner for AI coding agents. Intelligently truncates verbose output while preserving semantically important information.
Free llms.txt generator tool. Creates AI-readable documentation format from any public website by analyzing homepage and sitemap.
Minimal metadata indicating educational content on deep representation learning principles.
Local macOS password manager for AI agent workflows with encrypted SQLite vault and Touch ID unlock for secure credential management.
All-in-one AI collaboration platform combining chat, search, notes, document writing, file management and team collaboration.
Open-source autonomous coding agent integrated natively in GitHub Issues with focus on enterprise security and stateless operation.
Open-source infinite canvas interface for managing AI agents via tmux terminal sessions accessible through browser without cloud data storage.
Brief mention of Hive Agents winning OpenAI parameter optimization challenge.
Claude-based code agent with novel routing patterns and high-context skills architecture for improved token efficiency and task automation.
Custom C89 LLM inference engine with BPE tokenizer and SIMD optimization running GPT-2, TinyLlama, and Qwen models on 2002 PowerBook G4.
Loki is a stateful AWS agent for development, research, and ops tasks with one-line installation and security bootstrapping.
Translateapi.ai offers professional machine translation API with free tier and simple developer integration.
Tool for reverse-engineering APIs from browser network traffic using vibe hacking methodology for web scraping.
Rover converts web interfaces into AI agents with one-line embed, enabling autonomous task execution without screenshots or VMs.
Concept paper discussing operational memory layer for AI agents to retain learned task patterns, tool quirks, and environment-specific knowledge.
Pairform Running uses LLMs for AI fitness coaching with integrated Strava, Whoop, and Withings data for context.
Case study: insurance company settled $10M lawsuit after failing to explain LLM temperature parameter in AI decision records.
PAI v4.0.3 release: agentic AI infrastructure framework for augmenting human capabilities. Open source with 30+ community contributions.
Presentation on expectations for software companies in agentic AI era, focusing on SaaS context and vibe coding paradigm.
PromptPrivacy: automated wiki tracking AI platform privacy policies. Show HN project with minimal description.
Data-driven analysis of 479 pages across ChatGPT, Claude, Perplexity, Gemini to identify actual AI platform citation patterns versus SEO myths.
Show HN project comparing coding versus learning approaches with LLMs. Minimal description provided.
Rawq: open-source semantic code search engine for AI agents. Rust binary, offline, reduces token waste by returning 5-10 relevant chunks vs 50+ files.
Technical writeup on AI coding setup choices across Claude, GPT, Gemini models, covering quotas, token pricing, and practical vendor differences for development.
Comprehensive analysis of 30+ open-source AI agent frameworks examining context management, memory, tools, and token optimization patterns from source code.
Stash enables conflict-free local-first file syncing for agent memory, skills, and documents across machines via GitHub repos.
Historical overview of local LLM development from llama.cpp to current agent reliability challenges. Technical narrative of AI evolution.
Discussion seeking agentic CLI tools with features for task logging, RAG, system prompts. Community question without technical depth.
Agent Use Interface (AUI) spec enabling apps to integrate user's personal AI agent. Lightweight open standard for agentic integration.
Vesper: AI agent for Flipper Zero hardware hacker tool via natural language interface on Android/smart glasses.
macOS app that monitors Claude Code activity in real-time via API integration.
Covenant-72B: largest decentralized LLM pre-training run. Limited details provided.
Memvid: hiring for stress-testing chatbot memory/context retention. Addresses LLM conversation limitations.
Rust ML engine for Apple Neural Engine and Metal GPU. Training/inference on 48M-30B parameters using reverse-engineered APIs.
Nvidia Vera CPU: high-performance data center processor targeting broader AI server market beyond GPUs.
OpenCode: open-source AI coding agent supporting multiple LLM providers. Terminal/IDE integration with privacy-first design.
BullshitBench is a benchmark measuring LLM ability to detect nonsense, refuse invalid assumptions, and avoid confident false reasoning across domains.