A new workbench for running local AI models
Solo developer released Kivarro, an open-source local inference workbench for running AI models on personal hardware, built on Rust and Tauri and targeting GGUF models. The creator posted it to r/Loca…
Solo developer released Kivarro, an open-source local inference workbench for running AI models on personal hardware, built on Rust and Tauri and targeting GGUF models. The creator posted it to r/Loca…
Bridgewater's AIA Labs and Thinking Machines Lab fine-tuned a Qwen3-235B model on the hedge fund's internal investor judgement, achieving 84.7% accuracy on triage tasks versus 78.2% for the best front…
A community-built fine-tune of Google's Gemma 4 31B model has beaten the base model by 290 Elo points on the EqBench3 benchmark for marketing copy, according to a Reddit post. The fine-tune leverages …
NVIDIA turned its BioNeMo life-sciences stack into an agent toolkit, allowing AI agents to call accelerated models and libraries for scientific tasks. Anthropic's Claude Science workbench is the first…
Anthropic released Claude Sonnet 5 on 30 June 2026, claiming it nearly matches Opus 4.8 in agentic performance at a lower price. The model is the default on Free and Pro plans, with introductory prici…
Anthropic's Claude models became generally available on Microsoft Foundry, running on NVIDIA GB300 Blackwell Ultra GPUs, as part of a strategic partnership involving a $30 billion Azure compute commit…
OpenAI launched GPT-5.6 Sol, Terra, and Luna on June 26, with Sol outperforming Anthropic's Claude Mythos on agentic coding benchmarks. The US government restricted access to trusted partners under a …
Two new 2026 comparisons from DeepInfra and GMI Cloud conclude that the gap between open and closed LLMs has narrowed to 5-10% on overall capability, with no clean leaderboard existing. Closed models …
AWS published a technical how-to on June 25, 2026, detailing a pattern for adding agent-to-agent (A2A) capabilities to existing REST services without rewriting core code. The proposed overlay approach…
Retriever, a browser-agent startup, cut the cost of automated web workflows by over 100x by swapping its planning model from a frontier API to DeepSeek V4 Flash, an openly licensed Chinese model. A mu…
Anthropic launched Claude Tag, an always-on AI agent that joins Slack channels as a persistent team member, replacing its older Slack connector. The agent builds long-term memory of projects and decis…
Google's Gemma 4 31B outperforms Alibaba's Qwen 3.6 27B on agentic code review tasks, finishing faster due to superior Multi-Token Prediction (MTP) design, according to benchmarks and field reports. W…
Ai2 released Tmax-27B on 23 June 2026, an open-weight terminal-agent model built on Qwen3.6-27B that scores 43% on Terminal Bench 2.0 and 69% on TB Lite. The dense 27B model outperforms the sparse 397…
Earl Co released Sage Router, an open-source self-hosted gateway that exposes a single endpoint for AI agents to route requests to multiple model providers with automatic failover. The tool targets sm…
Developer Torgeir Helgevold fine-tuned a 600-million-parameter local LLM (Qwen 3:0.6B) to classify household questions into metadata categories, achieving 92% accuracy on a test set—up from 10% with p…
NVIDIA featured Eco Wave Power, an Israeli-Swedish wave energy company, in a blog post arguing that AI's next bottleneck is power infrastructure, not chips. The company's grid-connected wave stations,…
Google promoted its Interactions API to general availability on June 22, 2026, making it the default interface for Gemini models and agents. The new API introduces typed steps, managed agents with Lin…
A new open-source AI assistant stack combining the Hermes agent and MiniMax-M3 model on Nous Portal costs under £50 per month and can automate tasks like market briefings, inbox triage, lead research,…
Artificial Analysis released AA-Briefcase, a new agentic benchmark for long-horizon knowledge work, on June 18, 2026. Claude Fable 5 leads the leaderboard with 1587 Elo at $31 per task, while open-wei…
A two-year Claude subscriber outside the US had his account suspended on June 17, likely due to his use of the Fable 5 model, which Anthropic was ordered by the US government to disable for foreign na…