Skillgrade: "Unit tests" for your agent skills
Skillgrade, a new open-source tool, enables developers to create and run unit tests for AI agent skills, ensuring agents correctly discover and use custom skills. The tool supports multiple AI agents …
Skillgrade, a new open-source tool, enables developers to create and run unit tests for AI agent skills, ensuring agents correctly discover and use custom skills. The tool supports multiple AI agents …
A developer created a plugin for OpenCode that modifies chat headers to fix compatibility with GPT-5.6 Luna. The plugin sets the originator to 'codex_cli_rs' and the User-Agent to 'codex_cli_rs/0.0.0 …
A developer released Bunrun, an open-source local dashboard built with Bun and Svelte that lets users start, stop, and monitor development apps via a web UI. The tool delegates configuration to an AI …
A developer released E-- (English--), a programming language written in canonical English that compiles deterministically to Python, separating LLM-based code generation from runtime execution to ensu…
Local Motion, a new tool for VS Code and Cursor, lets developers run local coding agents with local LLMs on macOS, managing model loading and memory automatically. The tool selects compatible models, …
A developer created a diagnostic framework for SillyTavern Game Master character cards used in tabletop RPGs, identifying common failure modes such as mood-heavy personas, flat NPCs, and forced pacing…
Voicebox, a free and open-source AI voice studio, launches as a local-first alternative to ElevenLabs and WisprFlow, enabling voice cloning, speech generation in 23 languages, and dictation with full …
GenUI, a native Swift workspace for generative user interfaces, has been released as open-source software. The platform uses AI agents to produce declarative A2UI messages that are validated against a…
A developer has released aeovim, a Rust-based TUI that applies Neovim's modal editing model to orchestrate multiple LLM coding agents, wrapping Claude Code as child processes. The tool, already in dai…
ClickMinded's 'Little Guys Strike Back' webinar presented a blueprint for building an AI-powered editorial team using agents like Claude Code, Codex, or Cursor. The system assigns different models to …
A developer built an automated email response system using n8n, Gmail, and OpenAI's language model to speed up lead response times. The workflow triggers on new emails, uses an AI agent to analyze con…
A developer outlines rules for delegating tasks to subagents in a Claude-powered coding workflow, emphasizing context economy and reasoning quality over parallelism. The guidelines cover batch sizing,…
A developer tested an aggressive version of latent-space reasoning on a 1.5B-parameter language model, where the model pauses during generation to run parallel hidden-state rollouts without decoding t…
A developer released a real-time n-body simulation using the Barnes-Hut algorithm on GPU, implemented in CUDA C++ and OpenGL, scaling to millions of particles on an NVIDIA RTX 500 Ada Laptop GPU. The …
LoopVera, an open-source runtime evidence ledger for debugging vision pipelines and agent workflows, is being developed to provide a see-tweak-diff-sign-off loop without replacing existing operators. …
A developer argues that macOS's built-in CLI tools like grep, find, and cat are suboptimal for AI coding agents, recommending modern replacements such as ripgrep, fd, and jq for better context economy…
A solo developer built Prometheus, an autonomous research system that runs 24/7 on a single Linux workstation with one RTX 5090 GPU, generating its own questions and running over 130,000 experiments. …
Visionaire MCP v0.7 introduces a verification layer that lets AI coding agents visually inspect live web pages, identify the exact CSS rule, file, and line causing rendering bugs, and verify fixes aut…
A developer released a starter kit for running ONNX Runtime Web inside a Web Worker to keep browser-based machine learning inference off the UI thread, improving performance on mobile devices. The pro…
Snitch, an open-source deterministic prose claim verifier for AI coding agents, has been released. It monitors transcripts from tools like Cursor and Claude Code to flag false claims made by AI agents…