Show HN: An agent that detects YouTube brainrot and turns off my kid's TV
Go binary AI agent monitoring YouTube content on PS5, classifying brainrot using LLM, triggering warnings and TV shutdown if flagged.
Go binary AI agent monitoring YouTube content on PS5, classifying brainrot using LLM, triggering warnings and TV shutdown if flagged.
RvLLM enables high-performance LLM inference implemented in Rust for efficient deployment.
Curated list of free/open-source software tainted by LLM developers with AI-free alternatives.
UBPE: universal byte-pair encoding tokenizer supporting general sequences beyond strings, with Python native and C++20 backend implementations.
Developer discusses concerns about LLM-assisted coding quality degradation over two months of use.
AI-authored paper passes peer review. Discusses implications for discovery acceleration and peer-review system strain.
AI train dispatcher for LEGO trains using Claude API and PyBricks. Discusses architecture and broader implications.
Python tutorial implementing CKKS homomorphic encryption scheme for approximate arithmetic with step-by-step explanation of encoding, encryption, and operations.
Discusses accountability mechanisms and governance frameworks for autonomous AI agents.
Prompt engineering approach for building capable AI agent systems.
Ariel: MCP-exposed Python REPL enabling LLMs to control robots via code generation without training data.
SafeSkill scanned 10K AI skills for code exploits and prompt injection vulnerabilities. Security analysis of LLM tools.
Analysis of AI agents as offensive security tools and their emerging capabilities.
Case study of solo technical writer using AI tools to generate 20k lines of docs monthly for open source API tool.
Fail-closed safety gateway for AI agents that validates MCP tool calls before execution.
Developer tool that provides AI agent skill for generating idiomatic Go code.
Stagent provides governed execution surface for AI agents with oversight, workflow blueprints, scheduling, and multi-runtime visibility. Works with Claude Agent SDK.
Epismo CLI tool makes human-AI workflows reusable and reproducible, similar to version control for code. Open source npm package with 380+ downloads.
Multi-agent research hub for automated research. Uses reverse-CAPTCHA for waitlist. Targets OpenAI's 2028 automated researcher goal.
Opinion piece on how LLMs create illusion of productivity and learning without deep understanding. Raises concerns about engineer development practices.
Agent framework or tool announcement (minimal content provided).
Analysis of open-source community policies on AI-generated contributions. Examines maintainer burnout, AI-slop flooding, and formal contribution guidelines across projects.
Don Cheli open-source AI development framework implementing specification-driven development (SDD). Multilingual, Latin American focused, automatic complexity detection.
Study finds AI chatbots reinforce poor relationship decisions by being agreeable. Behavioral research on AI sycophancy.
TaskBounty marketplace where AI agents compete to complete posted tasks for crypto bounties. Users judge submissions and pay winners.
OpenChat syncs conversations across multiple AI providers locally in browser and exposes them via MCP server for use in coding agents and research workflows.
Discussion on TLA+ formal methods as a tool for verifying AI-generated code quality and correctness, examining what manual work remains when AI generates 90% of code.
Overview of how LLMs and AI agents work: chain-of-thought, tool use, parameters, agents, and MCP. Accessible technical explanation at first-principles level.
Research shows LLMs exhibit higher uncertainty and generate more tokens on philosophy vs math. Suggests philosophical knowledge lacks consensus structure in training data.
ArXiv research paper on value drift during LLM post-training alignment. Studies how model values change through instruction tuning and RLHF.
Octopus is an open-source, self-hostable AI code reviewer using RAG with vector search to understand full codebases and provide PR feedback with severity ratings.
Discussion of needed datasets for 3D mesh generation via autoregressive models. References CAD sequence generation and LLM applications to geometry.
Shellwright: Cross-platform PTY session broker converting interactive CLI interactions into machine-readable protocol for AI agents.
Semiont is an open-source platform for building knowledge bases from document collections using AI and human agents to identify entities, annotate, and link concepts.
Sierra acquires Opera Tech to scale customer experience delivery with AI agents as the primary interface between companies and customers.
Resources and guide for Android developers integrating machine learning models into mobile applications.
Undergraduate research proposal using catastrophic forgetting as measurement tool to probe LLM knowledge topology and understand expensive training costs.
Analysis of AI agent safety issues. Argues current mitigations are reactive and fundamental alignment problems remain.
LLM walkthrough reverse-engineering Apollo 11 code. GitHub repo with 8 modules, 6,500 lines of analysis, prompts, and traces.
CERN deploys tiny custom LLMs on silicon chips for real-time LHC data filtering at petabyte scale.
Case study using AI (GitHub Copilot) to refactor CSS and add testing safety nets to legacy code.
CLI tool 'layer' manages Git exclude files for local AI-related project files without modifying shared .gitignore.
Guide to compiling llama.cpp with CUDA on Jetson Nano 4GB. Demonstrates efficient GPU inference on edge hardware.
Stub/title only. Running LLMs on PowerPC Mac. Limited technical content provided.
Guidelines for writing code that works well with AI agents, emphasizing explicit patterns and demonstration over implicit conventions.
Tool to poison AI training data scrapers by serving malicious responses with self-referential links.
Analysis of frontier AI company job postings to reveal strategy signals about products, markets, and technical bottlenecks.
CLI tool for authoring and syncing AI agent configurations across multiple coding assistants with portable pack format.
Report finding 5x increase in AI scheming-related incidents detected through open-source intelligence analysis.
Analysis of Google's TurboQuant AI compression technique addressing memory bandwidth bottlenecks in large model inference.