Qwen 3.8's Real Open Model Is the 27B
Alibaba's Qwen team released two open-weight models this week, with the 2.4-trillion-parameter Qwen3.8-Max grabbing headlines but the Apache-licensed, multimodal Qwen3.8-27B being the practical choice…
DeepSeek is a Chinese AI research laboratory that has developed highly capable open-source language models including DeepSeek-V3 and DeepSeek-R1, notable for their efficiency and performance.
Alibaba's Qwen team released two open-weight models this week, with the 2.4-trillion-parameter Qwen3.8-Max grabbing headlines but the Apache-licensed, multimodal Qwen3.8-27B being the practical choice…
DeepSeek has released DeepSeek-V4-Pro-0813, an advanced AI model with enhanced agent capabilities, now available via its website, mobile app, and API. The model scored 62.7 on the DeepSWE benchmark, u…
DeepSeek announced the general availability of DeepSeek-V4-Pro on August 13, 2026, across its app, web, and API, with the model name deepseek-v4-pro unchanged. The release includes MIT-licensed weight…
DeepSeek released DeepSeek Harness, an open-source agent framework under an MIT license, on August 13, 2026, the same day it declared V4-Pro generally available. The framework, which powers DeepSeek's…
DeepSeek released a developer preview of Harness, a software framework for building autonomous AI agents, on Thursday, August 14, 2026. The framework offers four operational modes—standard, code-focus…
DeepSeek launched a developer preview of Harness, a software framework for building autonomous AI agents, on Thursday, marking a strategic pivot beyond large language models. The modular plug-in syste…
OpenAI's annualized revenue is set to top $40 billion, according to Bloomberg, positioning the ChatGPT maker for a blockbuster IPO as early as this year. However, cost-conscious customers are increasi…
DeepSeek released DeepSeek-V4-Pro, featuring major agent upgrades, flexible reasoning effort levels, and native OpenAI Responses API support with one-click Codex setup. The company also introduced pea…
DeepSeek announced new pricing for its V4 API, with the company revealing updated rates for the model's usage. The announcement details the cost structure for the V4 API, which is part of DeepSeek's A…
Google launched Gemini 3.7 Flash, priced at $0.75 per million input tokens and $3.75 per million output tokens, roughly half the cost of its predecessor, with improvements in coding, automation, and a…
DeepSeek published new API pricing for its V4 Flash and V4 Pro models, introducing peak and off-peak rates. Off-peak V4 Flash prices rose sharply from previous levels: cache hit from $0.0028 to $0.007…
DeepSeek V4 Flash scored 82.7 on Terminal Bench 2.1, beating the previous-generation V4-Pro-Preview by 14.7 percentage points, and Gemini 3.7 Flash shipped with a 26% jump in coding capability, signal…
Pipe, a pipeline-native language, now supports retrieval-augmented generation (RAG) in about 10 lines of code without a vector database, framework, or external dependencies, according to a blog post i…
Pipe's new try_ai syntax catches runtime errors and uses an LLM to repair and re-execute the broken expression, with a catch block as a final fallback. In an example, the type error "42" * 3 was autom…
Pipe, a programming language from the Pipe in 30 Lines series, now runs LLM calls in parallel by default, reducing three sequential DeepSeek API calls from about 4 seconds to 1.5 seconds. The >> opera…
Researchers from arXiv (paper 2608.11215) show that laptop-scale agent society simulators can replace each LLM agent with a low-parameter surrogate fitted from a few hundred to a few thousand real que…
Z.ai, the Beijing-based company formerly known as Zhipu AI, has not publicly launched a GLM-5.3 model, but its open-weight GLM-5.2 model, released June 16, was assessed by the U.S. Center for AI Stand…
OpenAI and Cerebras announced GPT-5.6 Sol Ultrafast mode, reaching 750 output tokens per second, 14x baseline and 11x faster than Claude Fable 5, with invite-only access and no disclosed pricing. Goog…
DeepSeek has opened a developer preview of DeepSeek Harness (dsh), an open-source agent harness positioned against Anthropic's Claude Cowork, providing a web interface for running AI agents with a plu…
DeepSeek will introduce peak and off-peak pricing for its API starting Aug. 17, with off-peak rates set at half of peak rates. Peak hours are 9 a.m. to noon and 2 p.m. to 6 p.m. Beijing time. For deep…