cd/entity/DeepSeek-R1· home entities DeepSeek-R1
grep -l @deepseek-r1 /news/*.json | wc -l → 53

DeepSeek-R1

mentions 53 type Organization page 1/3 feed RSS

// recent coverage 53 mentions

18:01
2026-08-27
pub.towardsai.net
artificial-intelligence

Your Context Length Decides What a Kernel Is Worth

A 2× faster attention kernel yields only 0.66% end-to-end speedup on a 1,024-token prompt with a 128-token answer under vLLM's --goodput ttft:500 tpot:50 promise, according to a seven-part technical s…

03:08
2026-08-22
byteiota.com
artificial-intelligence

GLM-5.3: Z.ai Hits Frontier Coding via Post-Training

Z.ai released GLM-5.3 on August 14, improving Terminal-Bench 3.0 coding scores from 4.6% to 28.3% solely through post-training, without changing the 743-billion-parameter mixture-of-experts architectu…

06:55
2026-08-11
research.nvidia.com
artificial-intelligence

Recursive Think-Answer Process for LLMs and VLMs

Researchers propose Recursive Think-Answer Process (R-TAP), a method enabling large language models (LLMs) and vision-language models (VLMs) to iteratively refine their answers, outperforming single-p…

20:48
2026-08-06
aiunderstanding.org
artificial-intelligence

Study Finds Frontier AI Models Split Under Steering Pressure

A preprint posted on August 6 found that six frontier language models—Claude Opus 4.7, GPT-5, Gemini 2.5 Pro, DeepSeek-R1, Qwen3.7-Max, and Llama-3.3-70B-Instruct-Turbo—split under steering pressure, …

00:00
2026-07-31
seangoedecke.com
artificial-intelligence

AI models need moral support to make discoveries

In 2026, AI models are producing a flood of mathematical discoveries, with prompting strategies as simple as asking for a breakthrough, according to software engineer Sean Goedecke. Goedecke argues th…

12:13
2026-07-28
byteiota.com
artificial-intelligence

$500 RL Fine-Tune Beats Claude Opus 4.6 on Real Task

Ramp and Prime Intellect published a case study showing a small RL-trained model, FastAsk, outperformed Claude Opus 4.6 by 4 percentage points on exact-match accuracy for financial spreadsheet retriev…

02:45
2026-07-28
discuss.huggingface.co
artificial-intelligence

The Next Intelligence Explosion Will Be Distributed

Three papers from Google, Harvard & MIT, and Imperial College & Huawei, all published in early 2026, converge on the conclusion that the next leap in AI will come not from larger monolithic models but…

19:46
2026-07-25
promptcube3.com
large-language-models

DeepSeek-R1 Local Deployment: My Hardware Struggles

A user reports that deploying the full 671B parameter DeepSeek-R1 model locally requires over 100GB of VRAM and is impractical on consumer hardware, with CUDA out-of-memory errors occurring even at sm…

12:11
2026-07-25
xn--vk5b17r.online
ai-policy

What if the RAM/GPU shortage is deliberate?

A theory suggests the RAM and GPU shortages felt by consumers in 2026 may be a deliberate side effect of AI companies hoarding hardware to prevent users from running free Chinese models like DeepSeek-…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics