cd/entity/DeepSeek-R1· home entities DeepSeek-R1
grep -l @deepseek-r1 /news/*.json | wc -l → 35

DeepSeek-R1

mentions 35 type Organization page 1/2 feed RSS

// recent coverage 35 mentions

12:11
2026-07-25
xn--vk5b17r.online
ai-policy

What if the RAM/GPU shortage is deliberate?

A theory suggests the RAM and GPU shortages felt by consumers in 2026 may be a deliberate side effect of AI companies hoarding hardware to prevent users from running free Chinese models like DeepSeek-…

15:00
2026-07-21
theregister.com
ai-infrastructure

Nvidia shows off Vera Rubin platform for tokenmaxxing

Nvidia demonstrated its Vera Rubin platform at a lab tour in Sunnyvale, California, claiming the compute tray can be assembled in one minute versus 90 minutes for prior GB200 hardware, a 90x improveme…

09:38
2026-07-21
blog.devgenius.io
large-language-models

Ollama Was Fun for About Two Weeks. Then Reality Showed Up.

Ollama, a tool for running large language models locally, initially impresses with ease of use but quickly reveals critical limitations for production use, according to a user account. The tool hides …

07:53
2026-07-21
snipvote.com
ai-safety

PlanFlip attacks achieve 0.68 success rate on GPT-5

A new arXiv paper reveals that PlanFlip attacks achieve a 0.68 attack success rate on GPT-5, with homogeneous multi-agent pipelines using the same LLM backbone for planning and auditing being highly v…

13:01
2026-07-20
pub.towardsai.net
machine-learning

GRPO from Scratch: Group Relative Policy Optimization in Python

Group Relative Policy Optimization (GRPO) eliminates the value model used in Proximal Policy Optimization (PPO), reducing active memory footprint from 42 GB to 28 GB for a 7B parameter model in FP16 b…

11:16
2026-07-18
magazine.sebastianraschka.com
large-language-models

Controlling Reasoning Effort in LLMs

OpenAI released the GPT-5.6 model family last week, which comes in three sizes each with roughly five or six reasoning-effort settings, according to Sebastian Raschka. The article explains how reasoni…

04:00
2026-07-14
arxiv.org
large-language-models

Reference-Based Distillation Detection in LLMs

Researchers at arXiv introduce a reference-based distillation detection method that identifies whether a large language model was distilled from a specific teacher model by comparing its output alignm…

19:00
2026-07-13
dev.to
large-language-models

Let Me Show You Which AI Model Actually Writes the Best Code

A developer benchmarked 10 large language models on five coding tasks, finding that DeepSeek V4 Flash offers the best value-to-quality ratio at $0.25 per million output tokens, while Qwen3-Coder-30B e…

22:04
2026-07-11
sourcefeed.dev
ai-infrastructure

Demystifying the NVIDIA DGX Spark for API Developers

NVIDIA's DGX Spark desktop GPU, with 128 GB unified memory and a 140W ARM64 processor, challenges API developers to shift from cloud-based AI consumption to local systems engineering. The device's sha…

11:00
2026-07-08
spectrum.ieee.org
large-language-models

AI Models Overthink Problems—and It’s a Security Risk

Researchers from Zhejiang University and Alibaba demonstrated a new denial-of-service attack on large language models by deliberately inducing "overthinking" through logically inconsistent prompts, ca…

17:32
2026-07-07
letsdatascience.com
large-language-models

EmulatRx demonstrates collaborative AI for clinical trial design

Weill Cornell Medicine researchers published EmulatRx in Nature Communications on July 7, 2026, a five-agent LLM framework for clinical-trial design using real-world patient data from EHR datasets suc…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics