cd/entity/DeepSeek V4· home entities DeepSeek V4
grep -l @deepseek v4 /news/*.json | wc -l → 59

DeepSeek V4

mentions 59 type Person page 1/3 feed RSS

// recent coverage 59 mentions

00:00
2026-09-10
mindstudio.ai
ai-products

How to Run Nex-N2.5 Mini Locally on RunPod (Dual H100 Setup)

Nex AGI's Nex-N2.5 Mini, the smaller model in the company's N2.5 agentic family, requires two 80GB-class GPUs such as H100s and consumed roughly 66GB of VRAM per card in testing on RunPod, according t…

10:02
2026-09-01
abliteration.ai
artificial-intelligence

Abliterated model large v2: GLM 5.3 84.5% CyberGym

Abliteration.ai released abliterated-model-large-v2, an abliterated version of GLM 5.3 hosted in FP8, scoring 84.5% pass@1 on CyberGym's 1,507 OSS-Fuzz bugs across 188 projects, 41.8% on Terminal-Benc…

15:46
2026-08-30
newsletter.semianalysis.com
ai-safety

Most Neoclouds Suck At Security

SemiAnalysis's ClusterMAX 3.0 testing found that most neoclouds have serious security vulnerabilities, with five frightening patterns identified. The company urges neocloud operators and users to upda…

09:52
2026-08-26
autorouter.top
ai-infrastructure

AutoRouter – Enterprise AI gateway for every model

AutoRouter, an enterprise AI gateway, now provides a unified API connecting to over 200 large language models from 20+ providers, including Seedance 2.0, Kling 3.0, GPT-5.5, Claude Opus 4.7, Gemini 3.…

13:30
2026-08-18
dev.to
large-language-models

Kimi K2 API Integration: A No-Fluff Getting Started

Moonshot AI's Kimi K2, a Mixture-of-Experts model, supports native image understanding through the standard chat-completions API, enabling multimodal tasks like document QA and screenshot analysis. Th…

00:00
2026-08-18
tomtunguz.com
artificial-intelligence

Birds Don't Fly Like Planes. Neither Does AI.

Qwen3.8-27B, a 27-billion-parameter dense model from Alibaba, ranks #1 of 135 models on Artificial Analysis's Intelligence Index with a score of 52, one point above GLM-5.2, a 753-billion-parameter op…

01:42
2026-08-16
blog.stackademic.com
large-language-models

I Gave DeepSeek V4 My Entire Codebase. Here’s What It Found.

DeepSeek V4, released as a preview on April 24, 2026, and fully available on July 20, 2026, offers a 1-million-token context window and pricing of $0.435 per million input tokens and $0.87 per million…

18:30
2026-08-14
lobu.ai
ai-agents

Self-Improving Agents Are Event-Sourced

Lobu, an AI startup, has built an event-sourced memory layer for AI agents that treats every write as an append-only commit, with corrections made by writing new events that supersede old ones, ensuri…

13:00
2026-08-11
developer.nvidia.com
artificial-intelligence

Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard

NVIDIA introduced NeMo Switchyard, a library that enables developers to route AI agent workloads across multiple models based on task requirements, cost, and latency, improving efficiency and accuracy…

18:31
2026-08-06
blog.kilo.ai
ai-tools

How Many Spoons Does Your AI Coding Tool Cost?

AI coding agents can drain a team's limited energy budget through vendor lock-in, broken environments, and silent failures, according to a piece on Kilo Code, an open-source coding agent. The model ma…

21:30
2026-08-04
dev.to
artificial-intelligence

How Much Does It Cost to Self-Host Open Models on AWS?

A developer analyzed the costs of self-hosting open-weights AI models on AWS, finding that a team of 10 can run a strong model like Llama 4 Maverick for about $1,250 per month during business hours us…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics