Anthropic Engineering Postmortem: Claude's 60-Minute Memory Bug
DeepSeek-V4 technical report: MoE language models with 1.6T/284B parameters, 1M token context, trained on 32T+ tokens with architectural improvements.
DeepSeek-V4 technical report: MoE language models with 1.6T/284B parameters, 1M token context, trained on 32T+ tokens with architectural improvements.
DeepSeek-V4 technical report: MoE language models with 1.6T/284B parameters, 1M token context, trained on 32T+ tokens with architectural improvements.
Genesis AI agent framework providing persistent memory, self-learning, and autonomous decision-making without constant user oversight.
AURA: Open-source agentic harness for SRE providing persistent orchestration and context management for production AI agents.
Verkor.io's Design Conductor AI agent autonomously designed complete RISC-V CPU core from 219-word specification in 12 hours.
Opinion on GPT-5.5 model capabilities claiming fewer tradeoffs than typical frontier models.
Project Omni: Multi-runtime cognitive system for developer workflows combining Rust, Python, and Node layers for agentic planning and execution.
Codex skill: Multi-agent system converting natural language prompts to game-ready 2D sprite sheets with image generation.
Sierra acquires Fragment, a YC-backed startup enabling AI workflow integration for enterprises. Third acquisition in customer service AI space.
llm.rb: Ruby runtime for building AI systems with agent and LLM capabilities, zero dependencies, Rails-compatible.
Fine-tuning Pi0.5 robot policy to 100% on LIBERO-Spatial and distilling it to single forward pass using SnapFlow for faster deployment.
Discussion: Solo developer balancing open-source AI infrastructure project exposure with algorithmic IP protection from LLM training.
MirrorNeuron: Open-source runtime for reliable on-device AI agent execution on edge hardware with improved software abstractions.
Selvedge tracks reasoning behind AI-generated code changes by capturing agent intent and context, addressing the problem of lost information when AI modifies codebases.
SparseLab: PyTorch sparse training library using CSR format and custom kernels optimized for CPU-first execution.
Overgrow plugin for Claude Code provides AI-powered SEO/GEO optimization with commands for auditing and generating landing pages.
Sakana Fugu is a multi-agent orchestration system coordinating frontier foundation models for coding, mathematics, and scientific reasoning tasks, now open for beta applications.
Analysis of enterprise AI agent security requirements, arguing MCP gateways need identity, authorization, and zero-trust mechanisms beyond routing.
Noscroll startup launches AI bot that monitors social feeds and news sites, sending summaries to users to reduce doomscrolling.
Newsletter commentary on Claude Opus 4.7 release, noting mixed reception regarding its intelligence, personality traits, and instruction-following behavior.
Article title only; discusses prompt injection attacks against AI systems on the web.
OpenAI releases GPT-5.5 model with two years of research, rolling out to ChatGPT and Codex users.
Technical analysis of NF4/FP4 4-bit quantization formats for LLMs, weight distribution properties.
CLI tool capturing coding agent sessions across teams, enabling shared learning and reducing duplicate investigations.
Open-source ML agent autonomously researches papers, trains and ships ML models using Hugging Face ecosystem.
Early access impressions of GPT-5.5, noting continued rapid AI improvement and remaining capability gaps.
Comparison of Jan.ai open-source alternative to LM Studio for running local LLMs.
Open-source hosting platform for AI agents supporting multiple content formats, includes MCP server and CLI.
Open-source JavaScript agent sandbox using object capabilities for secure AI agent execution with progressive permission granting.
Sophia: scalable second-order optimizer for language model pre-training with improved efficiency.
Nutanix announces agentic AI platform with GPU virtualization framework for multi-tenant enterprise deployment.
ArXivLean dataset evaluates LLM ability to formally prove mathematical theorems using Lean proof assistant, achieving perfect Putnam 2025 score by automatically verifying proofs as code.
PayClaw gasless USDC wallet library for AI agents across 12 frameworks. Open source tool for agent payments and transactions.
Jupyter notebook collection implementing ML algorithms from first principles with visualizations of training convergence.
Technical deep-dive on optimizing CUDA matrix multiplication kernels from 800ms to 25ms using harness engineering and human-in-the-loop workflow.
JetBrains survey of 10k developers revealing which AI coding tools are actually used in production work vs. hype.
Developer shares experience building frontend skills while working in ML/AI startups, discussing patterns noticed in AI outputs.
Farcaster Agent Kit CLI enables AI agents to interact with Farcaster social network via zero-cost hub protocol. Open source agent framework.
Blind benchmark testing 14 LLMs on WordPress plugin development after Copilot dropped Claude Opus support. Comparative evaluation of model capabilities.
Récif open-source platform deploys AI agents on Kubernetes with evaluation-driven lifecycle, governance, and monitoring. Production-ready agent infrastructure.
Natural language telemetry and intent analytics platform for AI products, evolved from monetization tool based on user demand.
TeamFuse: Open source template spinning up 5 Claude Code agents coordinating work via MCP-based messaging layer with Kanban integration.
Discussion on whether non-technical founders can build commercial products with AI, noting limitations of hands-off Claude Code usage.
Vision Banana achieves state-of-the-art on zero-shot transfer for 2D/3D vision tasks using image generators as generalist learners.
Thedex releases TinyThedex, a distilled model for log search that is 7x smaller, 5-7x faster with 96.4% quality retention.
Voxyflow AI tool that decomposes project descriptions into autonomous agent tasks, manages parallel execution, and maintains project context.
Tutorial on building multi-step agents in 50 lines of Python with critical perspective on LLM hype and real-world limitations.
Python tool that automatically adds notifications to long-running scripts without code modifications, useful for ML training and data processing tasks.
Prax agent runtime that iteratively fixes code errors. Website content description rather than technical documentation.
AgentSearch: self-hosted multi-engine search API with MCP server support for AI agents, no API keys or vendor lock-in.