deepseek-ai/DeepSeek-V4-Flash-0731
DeepSeek released DeepSeek-V4-Flash-0731, a model that Artificial Analysis ranks ahead of MiniMax M3, a 428B model, and at $0.14 per million input tokens and $0.27 per million output tokens it may be …
DeepSeek is a Chinese AI research laboratory that has developed highly capable open-source language models including DeepSeek-V3 and DeepSeek-R1, notable for their efficiency and performance.
DeepSeek released DeepSeek-V4-Flash-0731, a model that Artificial Analysis ranks ahead of MiniMax M3, a 428B model, and at $0.14 per million input tokens and $0.27 per million output tokens it may be …
DeepSeek's V4-Flash-High open-weight model scored 1586 on arena.ai's Frontend Code Arena Pareto Frontier, ranking #7 overall and #3 among open-weight models, a 154-point jump from the previous V4-Pro-…
DeepSeek released DeepSeek-V4-Flash-0731 on Hugging Face and moved its V4-Flash API to public beta on July 31, 2026, claiming major gains in agentic and coding benchmarks from re-post-training rather …
Manifest, an open-source development platform, shipped an LLM router in March 2024, saw 7,000 cloud users use it for four months, and deprecated it in June, with full shutdown on September 1. The comp…
Chinese AI startup Moonshot AI reportedly trained its 2.8 trillion parameter Kimi K3 model using Nvidia's high-end GB300 chips acquired in Thailand, circumventing US export controls, according to Whit…
OpenAI cut the API price of GPT-5.6 Luna by 80% on July 30, from $1.00 to $0.20 per million input tokens and from $6.00 to $1.20 per million output tokens, in a direct response to DeepSeek's launch of…
DeepSeek's V4 Flash 0731 model, priced at $0.26, achieved a score equal to OpenAI's GPT-5.6, which costs $5.01, on the Agentic Memory Benchmark, according to the leaderboard maintained by ATM Bench. T…
BarunLM, a 35-million-parameter language model introduced by developer Barun, outperforms models over 6x larger, including LiquidAI's lfm2.5-230m, while training on a single H200 GPU. The model's arch…
Chinese AI researchers are increasingly joining X to participate in global AI discussions, with Moonshot AI, the company behind Kimi K3, having about 30 affiliated accounts on the platform, including …
OpenAI has banned a cluster of ChatGPT accounts likely originating from China, dubbed 'Peer Review,' after detecting they were used to develop surveillance tools, analyze protest announcements, and re…
DeepSeek announced its API service will adopt a peak-valley pricing strategy, with peak-hour prices doubling the regular price across all billing items. Peak hours in UTC are 1:00–4:00 AM and 6:00–10:…
The European Union launched a new team on Friday to enforce its AI Act, which takes effect Sunday, requiring AI companies to label AI-generated content and comply with regulations addressing systemic …
The European Commission is adding 38 staff to its AI Office to enforce Article 50 of the EU AI Act, which takes effect on August 2, 2026, requiring companies to disclose AI-generated content, with fin…
China's State Council issued a new Regulation on Outbound Investment on June 1, 2026, effective July 1, 2026, targeting technology transfers and data flows involving AI, semiconductors, and green tech…
Moonshot AI, the Chinese lab behind the Kimi models, runs a large share of its training and serving on roughly 20,000 Nvidia H200 chips supplied through a compute agreement with Alibaba Cloud, accordi…
A security researcher detailed an autonomous AI attack that exploited exposed Langflow, n8n, and Marimo instances but was foiled when the agent left a public web server broadcasting its own tool calls…
A developer released a browser-based breakout clone designed for five-minute play sessions, featuring mouse or keyboard controls, cheats, and a physics engine decoupled from refresh rate running at 24…
Generative AI websites drew 9.5 billion visits a month worldwide between June 2025 and May 2026, up 70% year on year, while app downloads reached 4.4 billion, up 58%, according to Similarweb's 2026 Ge…
OpenAI cut the price of its GPT-5.6 Luna model by 80% on July 30, reducing input tokens to $0.20 per million and output tokens to $1.20 per million, down from $1 and $6, respectively, three weeks afte…
DeepSeek shipped the public beta of V4 Flash on July 31 with no architecture changes — same 284B MoE, 1M-token context — but a retraining-only round boosted Terminal Bench 25.8 points to 82.7, Toolath…