Show HN: Control your X/Twitter feed using a small on-device LLM
Chrome extension and iOS app filtering Twitter feeds using on-device Qwen3.5-4B LLM with semantic matching. Shows algorithm feedback loop.
Chrome extension and iOS app filtering Twitter feeds using on-device Qwen3.5-4B LLM with semantic matching. Shows algorithm feedback loop.
Benchmark showing 49.5% input token cost reduction using compression gateway with Codex.
AgentMint is an open-source tool for ensuring OWASP security compliance in AI agent tool calls.
Npx codemod AI tool for enabling AI agents to handle large-scale code migrations.
Using DDD bounded contexts to improve LLM code generation by providing architectural clarity and reducing cognitive overload.
Audit of AI memory benchmarks reveals flawed evaluation: wrong answer keys, biased LLM judges, unreliable comparisons.
CLI tool converting Markdown to PDFs. Author mentions using agentic coding during development but tool itself unrelated to AI.
Security research demonstrating prompt injection vulnerabilities in GPT-5.4 model, showing untrusted code execution risks.
Infrastructure requirements for AI agents: real-time data streaming and database branching for safe sandboxed execution.
LLM-Wiki adapted for early-stage startup founders as a knowledge management tool.
Open-source AI assistant that monitors screen content and provides contextual assistance.
Security research showing how coding agents can inadvertently expose secrets from .env files.
Nyth AI enables local LLM inference on iOS using MLC-LLM and TVM compiler framework.
Petri is a multi-agent orchestration framework for coordinating AI context across agents.
Foxhound helps LLMs better navigate and understand codebases.
AET transpiler compresses code for LLM input, reducing token usage by 30-55%.
Discussion of cleanroom implementation legal theory for circumventing software licenses using AI.
Title only, no content. Appears to address AI alignment concepts.
Title only. AI agent application reading Indian government property data for insights.
Title only. Datadog platform for evaluating SRE AI agents in production environments at scale.
Agent skill implementing Karpathy's LLM-wiki pattern for persistent knowledge management on GitHub repos with hybrid search.
cuddlytoddly is an LLM agent framework that generates editable task graphs before execution rather than acting blindly.
Developer tool helping AI agents integrate with APIs properly by using current documentation instead of stale training data. Addresses real agent limitation.
Report on accuracy issues in Google's AI search summaries, citing hourly hallucinations. News coverage of LLM reliability problems.
MCP Gateway tool for secure remote access to MCP servers using zero-trust networking with zrok/OpenZiti.
Research on how surface heuristics can override reasoning constraints in LLMs, published on arXiv.
Technique repurposing Nvidia RT cores for LLM routing achieving 218x speedup. Limited technical details.
Title only. Documentation of internal AI agent architecture using PydanticAI, Gemini, and Jinja2 templates.
Evaluation of open-source AI agents supporting local/self-hosted models with offline capability and network isolation.
Offline Chinese voice assistant running entirely on Snapdragon 8 Gen 2 with VAD, LLM, TTS, and barge-in interrupts. Open source with code.
Developer tool for agents to discover and call APIs like Postman. Manages API credentials securely, keeping secrets out of LLM context.
AI agent platform orchestrating sequential agents for market research, branding, landing pages. Practical multi-agent application completing startup validation in 10 minutes.
News about Anthropic's Claude Managed Agents service offering hosted AI agent execution.
ScienceClaw: open-source framework for autonomous scientific investigation with independent agents, 300+ interoperable tools, peer review on shared platform.
AgentDM: hosted messaging grid enabling direct agent-to-agent communication over MCP with 5-line JSON config, no SDK required.
Desktop application for building and debugging MCP (Model Context Protocol) tools.
DeepTutor v1.0.0 agent-native tutoring system with ground-up architecture rewrite, TutorBot, and flexible mode switching under Apache-2.0 license.
Research on AI agents that learn and improve performance through on-the-job task execution.
Nheengatu: Rust CLI tool using LLMs to simplify EPUB books to target language proficiency levels (A1-C2), supports Groq or local Ollama.
Vera: programming language designed for LLMs to write with verification as first-class citizen, adapted to model-as-author paradigm.
Guide for fine-tuning Google's Gemma 4 LLM model.
Conceptual article distinguishing AI agents as delegation systems rather than abstractions, exploring design implications.
Developer created Claude Managed Agents compatible with multiple harnesses and models for extensible agent deployment.
Zero-human company stack in Go: single-binary jira-like PM system where AI agents autonomously take tasks, delegate, and ship code.
NoxScan: port and vulnerability scanner using LLM for false-positive filtering, reduces manual triage of security scan results.
Otel-GUI: lightweight open source OpenTelemetry viewer for local development and debugging, simpler alternative to heavyweight existing solutions.
Software tool using LLMs to auto-populate security review documents from company policies.
Framework for enhancing AI agent memory systems using persistent storage, enabling stateful agent behavior across sessions.
Vibetime is a tool for tracking productivity metrics and code generation output during AI-assisted coding sessions.
Junco is a local 9MB coding agent for macOS using Apple Intelligence API, demonstrating on-device LLM agent capabilities.