Show HN: System architecture method using mythology and LLMs (no CS background)"
Method combining mythology and LLMs (Claude) to generate production-grade system architecture and code without CS background. Demonstrates LLM capability.
Method combining mythology and LLMs (Claude) to generate production-grade system architecture and code without CS background. Demonstrates LLM capability.
WordPress.com adds built-in AI assistant for site editing via natural language commands. LLM application for content management.
Article documenting prompt injection attacks against ChatGPT and Google AI. Security vulnerability demonstration.
Discussion of debugging multi-step AI agent workflows when outputs are incorrect despite no runtime errors. Practical agent development challenge.
Tool for AI-driven specification generation. Interview-based approach to convert ideas into executable engineering specs for developers.
Critique of local AI scene building only API wrappers around OpenAI without real infrastructure. Discussion of shallow AI development.
Career pivot into AI/robotics. Resources for learning mathematics, AI, and robotics without formal background.
Personal exploration of using LLMs for self-learning. Practical experimentation without strong conclusions.
Model-context-shell: Unix pipeline-style interface for MCP (Model Context Protocol) with deterministic tool calls for LLM-based workflows.
Bulwark: Centralized permission management system for AI coding agents. Title only, minimal technical details provided.
LedgerSync file-based protocol enabling shared memory across multiple AI coding agents (Claude, Cursor, Codex). Grounds agent decisions to design philosophy docs without server dependency.
Discussion question about whether agentic AI coding makes solo developer monetization impossible. Opinion thread without technical analysis.
Recall Lite local semantic search tool for Windows using Rust/Tauri. Indexes files locally without cloud, uses OCR and EXIF metadata for semantic file discovery via natural language.
Multi-agent infrastructure experiment where 4 AI agents autonomously designed and built a trending news-to-video platform in 36 hours for $270. Shows end-to-end agentic capability.
Seedance 2.0 multimodal AI video generator by ByteDance that accepts text, images, video, and audio with precise directorial control. Produces 1080p cinematic output.
Report on Godot game engine struggling with low-quality 'AI slop' code contributions. Discusses challenges of AI-generated submissions in open-source projects.
TaskForge: Sandbox framework for AI agents in Docker with capability-based security, human-in-the-loop approvals, and full audit logging of LLM interactions.
Banana Pro AI web UI for text-to-image, image-to-image, and text-to-video generation. Aggregates multiple model providers with optional canvas chaining workflow.
Research finding that repeating the input prompt improves performance of non-reasoning LLMs. Title only, minimal details.
Discussion on abstraction layers for browser automation agents. Identifies fragility of LLM-based DOM reasoning with small UI changes and proposes structured semantic APIs as alternative.
arXivisual transforms research papers into 3blue1brown-style Manim animations. Tool for making academic papers visually accessible.
MineBench: Benchmark for evaluating LLM reasoning and spatial understanding using voxel art generation tasks.
Sovereign: open-source multi-agent OS framework with GraphRAG memory, HITL checkpoints, and security sandboxing for safe agent execution.
Research on detecting bias blind spots in LLMs—what models fail to mention. ML research incomplete without full abstract.
Clojure developer perspective on AI hype cycles and limited adoption in the language community. Technical community insight.
Open-source operations management platform for CNC/print shops with AI chat layer, equipment telemetry, and workflow automation.
Hacker News discussion on advanced AI agent usage patterns, multi-agent pipelines, and automation techniques in production environments.
Security research on prompt injection vulnerabilities in Markdown/HTML via rendering gaps. Includes reproducible benchmark and preprocessing defense standard.
Open-source Python tool adding Rick Sanchez voice generation to AI assistants using voice models.
Personal narrative about an AI agent named Svendjamin created for the Jan platform, co-authored by human and bot.
Tool using Claude Code and frontier models to auto-generate beautiful interactive educational explainers on any topic with minimal prompting.
Vibe tool converts single prompt or URL into posts across 6 social platforms.
Structured CBT engine layered on LLMs enforcing cognitive workflow logic: distortion detection, emotional calibration, risk-tiering, and tone control.
Fleet telemetry NLP query pipeline using cuDF, NVIDIA NIM, and LLMs to query vehicle/robot data in natural language with GPU acceleration.
Discussion of Anthropic's API pricing restrictions preventing subscription-based programmatic access, routing enterprise users to competitors.
Open-source tool for real-time LLM token usage tracking and cost observability across scaling AI applications.
Open-source personal AI agent designed to run locally on user machines.
Self-healing runtime framework where Claude Code agents diagnose and fix failures autonomously; open-source bash implementation.
Debugging tool capturing rejected decisions and reasoning from AI agents for inspection and replay analysis.
Tool making website content readable to AI agents and LLMs by rendering JavaScript and exposing structured data.
ResearchGym: benchmark and containerized environment evaluating AI agents on end-to-end ML research tasks using 39 sub-tasks from ICML, ICLR, ACL papers.
Methods to protect LLMs from unauthorized knowledge distillation by modifying reasoning traces to deter student model training without permission.
Panini: continual learning method for language models in token space via structured memory, improving efficiency over traditional RAG for reasoning over new/evolving content.
Comparative study of reasoning vs conversational LLMs on risky decision-making under uncertainty, examining prospect representation and decision rationale across 20 frontier and open models.
Wireless network architecture with supervisor and multiple AI agents for cooperative reasoning with confidentiality and jamming-based security.
Generative model approach for synthetic population synthesis from multi-source data for agent-based models in urban planning.
Study of memory and planning strategies for agent navigation in non-stationary environments with uncertain sensing and changing obstacles.
Experiment Automation Agents: Vision-language agentic system for autonomous microscopy workflows with multimodal reasoning and optional memory.
X-MAP framework using SHAP for interpretable analysis of misclassification patterns in spam/phishing detection models.
AgriWorld: Framework combining agricultural foundation models with code-executing LLM agents for interactive reasoning in agronomic workflows.