cd/entity/DeepSeek· home entities DeepSeek
grep -l @deepseek /news/*.json | wc -l → 1696

DeepSeek

DeepSeek is a Chinese AI research laboratory that has developed highly capable open-source language models including DeepSeek-V3 and DeepSeek-R1, notable for their efficiency and performance.

mentions 1696 type Organization page 83/85 feed RSS
sameAs · en.wikipedia.org · www.wikidata.org

// recent coverage 1696 mentions

12:31
2026-05-27
github.com
large-language-models

I ran GLM-5.1 on a 16GB RAM machine

A team of engineers successfully ran the 754-billion parameter GLM-5.1 large language model on a consumer PC with only 16GB of RAM and a Ryzen 5 5600G CPU, achieving zero crashes or out-of-memory erro…

07:02
2026-05-27
letsdatascience.com
ai-agents

Pullfrog Delivers Open-Source GitHub Actions Automation

Pullfrog, an open-source AI-powered GitHub bot created by Colin McDonnell, launched in beta on May 12, 2026, running entirely within GitHub Actions. The bot uses a model-agnostic, bring-your-own-key a…

21:49
2026-05-26
letsdatascience.com
artificial-intelligence

China Restricts Overseas Travel for Top AI Talent

Chinese government agencies have begun requiring top artificial-intelligence professionals at private firms, including Alibaba and DeepSeek, to obtain approval before traveling abroad, according to pe…

14:25
2026-05-26
mayberay.bearblog.dev
large-language-models

Model is currently experiencing high demand

A developer's side-project, Ikka, which relies on the Gemini API for news summarization and ranking, has been repeatedly disrupted by a "high demand" error message, sometimes lasting for days. The dev…

07:42
2026-05-26
djdumpling.github.io
machine-learning

paper reading catalog

DeepSeek researchers introduced manifold-constrained hyper-connections to restore the identity mapping property in transformer architectures, addressing training instability and scalability issues cau…

06:39
2026-05-26
djdumpling.github.io
large-language-models

Frontier Model Training Methodologies

Seven open-weight frontier models, including Hugging Face's SmolLM3, DeepSeek-R1, and OpenAI's gpt-oss-120b, were analyzed to distill common training methodologies for multi-billion parameter models, …

04:00
2026-05-26
arxiv.org
large-language-models

BODHI: Precise OS Kernel Specification Inference

Researchers have developed BODHI, a domain knowledge prompting method that improves automated generation of formal operating system kernel specifications using large language models. The technique, wh…

02:18
2026-05-26
uumuse.ai
ai-products

Cited AI Workspace: No More Re-Uploading Files

UUMuse launched a cloud-hosted knowledge base that lets users upload files once and access them across multiple AI models, including GPT, Claude, DeepSeek, and Qwen, without re-uploading. The platform…

← prev page 83 / 85 next →
// co-occurs with top 8 entities
// topics top 6 topics