cd/entity/DeepSeek· home› entities› DeepSeek
grep -l @deepseek /news/*.json | wc -l → 2242

DeepSeek

DeepSeek is a Chinese AI research laboratory that has developed highly capable open-source language models including DeepSeek-V3 and DeepSeek-R1, notable for their efficiency and performance.

mentions 2242 type Organization page 12/113 feed RSS
sameAs · en.wikipedia.org · wikidata.org

// recent coverage 2242 mentions

12:00
2026-09-14
kdnuggets.com
artificial-intelligence

Why DeepSeek-V4.1-Flash Is Such an Exciting Open Model Release

DeepSeek released DeepSeek-V4.1-Flash, a 552B-parameter Mixture-of-Experts open model under an MIT license that activates only 8B parameters per token during prefill and 16B during decoding, supports …

09:57
2026-09-14
ziggit.dev
ai-tools

What's everybody working on? (September Edition)

A developer known as zigster64 published atari.zig, a repository that enables an LLVM m68k backend to produce a bare-bones Zig toolchain on macOS, and used DeepSeek tokens to have an LLM generate a Sp…

07:04
2026-09-14
news.slashdot.org
artificial-intelligence

Should US Open-Weight AI Labs 'Distill' Frontier Models Too?

Y Combinator CEO Garry Tan told TechCrunch he would "do nothing" to curb AI model distillation and instead wants U.S. regulators to let smaller American open-weight labs distill from domestic frontier…

06:38
2026-09-14
trackllm.net
ai-infrastructure

Show HN: TrackLLM: Are the LLM APIs you rely on stable?

TrackLLM, a monitoring project posted to Hacker News, reports that 839 of 1,743 cataloged LLM API endpoints are being tracked and that 191 changes occurred across 150 endpoints, with probe-prompt outp…

06:08
2026-09-14
byteiota.com
ai-agents

OpenCode: The Open-Source AI Coding Agent You Own

OpenCode, the MIT-licensed terminal coding agent built by SST (Anomaly Innovations), reached 195,000 GitHub stars and 16 million monthly users by September 2026 after Anthropic invalidated the OAuth t…

03:36
2026-09-14
ianbarber.blog
large-language-models

Agents love prefill

DeepSeek's V4.1 Flash technical report introduces a "Causal Encoder–Decoder" architecture that runs only the first 20 of the model's 40 layers during prefill, cutting prefill compute in half for its 5…

21:12
2026-09-13
dev.to
large-language-models

Your LLM cost estimate is wrong above 200,000 tokens

A developer's survey of LLM API pricing found that most comparison tables understate costs for long-context workloads because several providers double their per-token rates above a context threshold. …

← prev page 12 / 113 next →
// co-occurs with top 8 entities
// topics top 6 topics