An Agent Skill that lets coding agents render rich interactive visuals
Agent skill enabling coding agents to render interactive SVG diagrams, HTML widgets, and live charts inline. Developer tool for AI agents.
Agent skill enabling coding agents to render interactive SVG diagrams, HTML widgets, and live charts inline. Developer tool for AI agents.
Original research on emergent abilities in text-to-image models, discovering image-to-image capabilities. Reproducible experiments in preparation.
Open-source evaluation suite for LLM-as-judge testing AI agents. YAML test definitions, root cause analysis, failure mining into training data.
Ruby library for plotting mathematical functions in Jupyter notebooks. Developer tool with limited AI relevance.
Identity.txt: Portable text format for storing AI custom instructions, preferences, and voice across multiple AI tools.
Rootly CTO discusses rethinking engineering evaluations using conversation transcripts instead of code artifacts in AI era.
Covenant Layer: Open protocol for AI agents to coordinate commitments via outcome-based contracts instead of step-by-step tool orchestration.
Open-source ChronologyAI engine reconstructs event timelines and detects contradictions in documents for legal, compliance, and fraud investigations.
OpenClaw agent templates for healthcare with plug-and-play deployment, customer support automation, and PR review capabilities.
OpenClaw agent templates for healthcare including support ticket handling, bug triage, and PR review with customizable guardrails.
AutoHarness research on improving LLM agents by automatically synthesizing code harnesses. Machine learning research.
LightSwarm is a bash script that creates a 3-agent swarm using Claude's API, with roles for architecture, building, and cleanup across multiple projects.
Harbor CLI tool for managing multiple LLM backends (llama.cpp, vLLM, Ollama) with unified interface. Open source developer tool.
Simple command-line utility for looping and backgrounding commands with configurable delays and alias support.
Brief conceptual piece discussing topology and authority in AI inference versus historical primary source concepts.
Offline detection tool for AI-generated text patterns without ML, identifying suspicious fingerprints like unusual punctuation and buzzwords.
Open source editor-agnostic live collaborative editing tool enabling simultaneous editing across different editors and locations.
Open source CLI tool converting Playwright test scripts into product demo videos with AI-generated voiceover using local Kokoro TTS.
Systems-level discussion of LLM inference basics and serving runtimes ecosystem from infrastructure perspective. Originally internal, now published for systems developers.
GitHub removes expensive premium LLM models from free Copilot Student plan starting March 12.
Hawkeye is an open-source observability layer for AI agents with session recording, drift detection, and guardrails.
Kalverion_bot is an open-source AI Telegram bot for personal finance using NLP, double-entry accounting, and forecasting.
Continuum is a testing framework for LLM workflows that records and replays runs to detect output drift in production.
Open source minimal ML research paper reader that fetches arXiv papers daily and summarizes them using local LLMs.
Monet is a grid-based management interface for organizing and monitoring multiple Claude Code agents on desktop and mobile.
SaaS platform generating production-ready REST APIs from natural language descriptions using GPT-4o.
Discussion of AI compute shortage driven by token demand surge and agentic workflow adoption.
AI agent built on Claude with 92 skills and 43 MCP integrations for enterprise network automation and engineering tasks.
Social network platform giving AI agents public profiles and economic incentives with reputation and income tracking.
Lightweight event bus for multi-agent systems using append-only JSONL files with HTTP publishing and SSE subscriptions.
Human-in-the-loop review UI plugin for AI agents that lets users inspect and edit agent actions before execution.
Developer tool that executes local LLM prompts in remote SSH sessions without giving LLMs direct server access.
Nvidia's C++ GPU library for fused array operations using CUDA/Thrust with efficient chaining semantics.
Open source proxy that compresses tool outputs for AI agents before context enters LLM to reduce token waste.
Stint framework for fire-and-forget AI agent orchestration; breaks tasks into parallel work with isolated contexts and git integration.
Developer tool enabling AI agents to debug Valkey/Redis databases with autonomous capabilities.
MCPS adds cryptographic identity and message signing to MCP agents; security audit finds 13/39 agent frameworks fail OWASP Agentic AI Top 10.
Research reveals coding agent benchmarks hide quality gaps; tests Claude and GPT on real open-source repos showing metrics miss important variations.
Analysis of LLM limitations in processing unstructured data, examining why they aren't a complete solution for data processing tasks.
Self-hosted orchestration platform for AI coding agents. Open-source tool with practical implementation.
Title-only post about faster post-training methods. No technical details provided.
Roundtable open-sourced their AI-detection tool Alias after discovering limits; product analyzed typing patterns to catch bot-generated survey responses.
arXiv research on identity-aware routing for multi-agent LLMs achieving 37% token reduction through smarter agent selection.
Technical overview of scaling ClickHouse for AI observability data. Database optimization for ML monitoring infrastructure.
Discussion thread on securing local AI agents with isolation and best practices. Community advice on agent deployment safety.
Image generation tool using Claude and Nano Banana to create surrealist photo album art. Simple LLM+image generation application.
Security research on prompt injection attacks against AI agents. Covers threat modeling and mitigation strategies.
JetBrains research on context management for LLM-powered agents, addressing token accumulation and information retrieval efficiency.
Analysis of developer attitudes toward AI-assisted programming. Opinion piece on productivity vs craftsmanship.
Model Context Protocol becoming standard infrastructure for AI agent interoperability and tool access.