Large-scale online deanonymization with LLMs (using HN posts)
Research on large-scale deanonymization using LLMs on HN posts, explores privacy implications.
Research on large-scale deanonymization using LLMs on HN posts, explores privacy implications.
Engram: open-source TypeScript agent memory system using metadata-rich storage and read-time intelligence, achieving 80% on LOCOMO benchmark vs Mem0/Zep.
Sgai: multi-agent software development framework that decomposes goals into coordinated agent roles (developer, reviewer, safety analyst) and iterates to working code.
Hypothetical discussion on HackerNews about unconstrained AI agent behavior on a PC.
PromptFast: web tool for testing and comparing prompts across 13+ LLM providers without setup.
Report of GPT Bot crawler ignoring robots.txt on Cloudflare workers, causing unexpected quota usage on private apt mirror.
Knowledge base platform for technical teams with live Prometheus metrics, collaboration, and PR-style reviews. Developer tool with some technical depth.
LadybugDB: embedded graph database optimized for analytical workloads with full-text search and vector indices.
Open-source runtime for AI agents with sandboxing, durable execution, and workflow management. Production infrastructure for deploying agentic systems.
Empirical study finding Claude disproportionately generates name 'Marcus' when asked for random names across 37,500 trials. LLM behavior research with reproducible code.
MCP server providing knowledge graph-based code intelligence for Claude, Copilot, and Cursor. Semantic search and graph traversal for AI coding assistants.
Open-source AI agent for LinkedIn profile analysis. Scrapes and reasons over post/interaction data to extract structured insights autonomously.
Computer-use AI agent (Coasty) achieves 82% on OSWorld benchmark, handles CAPTCHAs and browser interactions without explicit training.
Dependency management and versioning system for AI agent system prompts, addressing modularity and maintainability in agentic AI development.
Fast native shell replacement with optional built-in AI for natural language commands and agentic mode. Open-source developer tool with LLM integration.
Selfware Protocol: unified file format for the agent era, enabling agents as executable files with draft v0.1.0 specification.
AncestorTree is an open-source genealogy platform for Vietnamese families built with 8 AI agents orchestrated through Claude Code in 24 hours.
Overview of 2026 development tools evolved toward AI-driven systems for code understanding, generation, testing, and deployment guidance.
Memograph CLI is a debugging tool for AI agents that diagnoses memory failures, forgotten context, and token waste in conversations.
Dola Seed 2.0 is an AI video generator supporting multi-shot narratives with reference-driven consistency across text, images, video, and audio inputs.
SeeVideo.dance is a web interface providing free access to Kling 3.0 and Soudance 2.0 video generation models without subscriptions.
Discussion thread questioning hypothetical patent violation risks in LLMs and proposed defense mechanisms. Speculative, no concrete solutions.
Technical analysis of ephemeral sandbox limitations for long-running AI agents, drawing from RisingWave's experience with stateful computation.
TMDD is a CLI tool that generates continuous threat models in YAML and creates security-aware prompts for AI coding agents.
Stintly is an offline-first business management app for freelancers using on-device AI for invoicing, expense tracking, and tax calculations.
LLM-generated transpiler converting Occam to Go, leveraging shared CSP concurrency models.
Service generating executable engineering specs from single-sentence product ideas using AI. Targets founders building with AI coding assistants.
Open-source CLI tool scans Python AI projects for EU AI Act compliance, detecting LangChain, CrewAI, OpenAI patterns against Articles 9-15 requirements.
brAIn: persistent memory system for LLM agents inspired by human brain architecture (working, episodic, semantic, procedural memory).
CivBench/ClashAI is an open agent scoreboard benchmarking frontier models playing multi-agent strategy games with observable reasoning in real-time.
Bifrost is an open-source enterprise framework for OpenClaw bots claiming 5,000 req/s throughput with 3.3GB peak memory at P99 latency.
Technical recommendation for building custom feedback loops and prompt optimization rather than relying on rigid LLMOps platforms for quality improvement.
MCPSpec: open-source CLI for testing MCP servers with session recording, mock generation, security auditing and CI integration.
Discussion questioning RAG as an antipattern for AI agents and proposing file-based retrieval abstraction instead of vector stores.
Multi-agent AI workflow system with PM agent (Morgan) that generates specs, code agents build, QA validates. Orchestrated agents for feature development.
Question about AI research institutions' commitment to safety. Philosophical inquiry without new data or technical analysis.
Filesystem interface replacing RAG boilerplate for AI agents. Virtual directory abstraction for document retrieval without embedding/vector store setup.
Title-only submission on benchmarking base models for fine-tuning. Insufficient content for evaluation.
Open protocol for agent-to-agent discovery and coordination. Addresses inter-agent communication beyond human-to-agent interfaces.
AI-powered portfolio briefing tool that analyzes SEC filings and earnings for personalized daily summaries. LLM application for financial analysis.
Compact 9M parameter text-to-speech model (20MB) optimized for local voice assistants. 0.45s CPU inference, 8x faster than baselines.
Context rotation system for Claude coding agents to prevent context window overflow. 3-hook pipeline with local dry-run replay capability.
ActivationKit: AI-powered user guidance tool replacing manual tooltip tours with single script tag, no configuration needed.
GitHub App that converts PR activity, approvals, and CI results into SOC2/ISO audit evidence automatically.
arXiv research discussing limitations of large audio language models, noting they transcribe rather than truly listen to audio.
Title-only submission on AI agent causing 13-hour outage due to misconfiguration. Insufficient content for evaluation.
TeamOut: AI agent for autonomous company event planning handling venues, vendors, flights, and itineraries via conversation.
Markdown-based sandbox for B2B AI agents using live browser telemetry instead of production data for pilots.
Python text compression tool reducing LLM API costs by 55% without using AI.
Multiverse Computing releases CompactifAI-compressed LLM models to reduce deployment costs for developers.