Sentinels: The Quiet Power of a Touched File
A developer uses sentinel files—empty files on disk—to gate risky actions by coding agents, such as exiting plan mode or opening pull requests. The sentinel acts as a durable state outside the agent's…
A developer uses sentinel files—empty files on disk—to gate risky actions by coding agents, such as exiting plan mode or opening pull requests. The sentinel acts as a durable state outside the agent's…
A developer named Chris, who has relied on AI coding tools like Claude Code and Cursor for over two years, argues that writing has become a critical skill for escaping "AI delirium." He explains that …
A developer used Claude Code with Opus 4.8 and its new workflows feature to autonomously build a complete incremental reading application in about 16 hours, consuming $1,000 in tokens and 46% of a Max…
A software developer using Claude Code reports writing significantly less code while spending more time understanding and testing AI-generated code, describing the shift as a positive change that pres…
NVIDIA's DGX Spark single-node AI appliance can run OpenAI's open-weight GPT-OSS-120B sparse MoE model at ~50 tokens/s using optimized engines like SGLang or llama.cpp, making it viable as a local cod…
A developer has released a free cloud-based tool that allows users to manage AI agents across multiple hosts from a single workspace. The tool connects agents like Claude Code and Cursor via a lightwe…
A senior executive at a major public tech company told the author that nearly all of her 1,000 engineers use Claude Code, yet individual productivity gains are not translating into proportional organi…
Anthropic released the Claude Agent SDK, providing the same engine that powers Claude Code as a programmable tool for developers. The SDK allows users to build custom terminal interfaces in about 10 m…
A single Claude Code session logged approximately 1,270 model turns and cost $1,278, with two-thirds of that bill — roughly $843 — going to re-sending context the model had already seen on every turn.…
Overslash, an open-source authentication gateway for AI agents, launched today as a centralized control point for managing agent permissions across services like GitHub and AWS. The tool allows develo…
A new Claude Code plugin called Claude HUD provides real-time visibility into context usage, active tools, running agents, and task progress directly within the terminal. The plugin installs via Claud…
A developer discovered that having ten Claude Code plugins active simultaneously was silently draining their €200 monthly plan's credits twice as fast as expected. Each active plugin injects roughly 2…
OpenAI integrated Codex into the ChatGPT mobile app on May 14, 2026, providing an official first-party path for mobile code development. Five realistic options now exist for using Codex from a phone, …
A new architectural paradigm called the "agentic mesh" proposes embedding AI agents into business processes to enable cognitive automation at scale, borrowing principles from the failed data mesh appr…
AI coding agents operate as a simple control loop: they build a message bundle, send it to an LLM, execute any requested tool calls, and repeat until the task is done. The LLM never directly executes …
AnyFrame launched a platform that lets teams build and deploy AI agents for any use case, integrating with existing tools like Slack, Claude Code, and Cursor. The platform provides a Python SDK and ru…
Vectoralix launched a new platform that converts Git repositories into hosted MCP (Model Context Protocol) servers, eliminating the need for developers to build custom protocol handling, authenticatio…
A photographer has built AgentFlow, a declarative DSL that lets users define multi-agent AI workflows in simple `.aflow` files and compile them into MCP tools for Claude Code without writing any integ…
Komi-learn, a new tool for coding agents, automatically learns a user's coding style, stack, and fixes from sessions and recalls them in future sessions without commands. The tool works with Claude Co…
A developer in Tokyo conducted a 30-day structured benchmark of four Claude models in Claude Code, tracking token usage, response quality, and cost per task type. The results contradict the consensus …