cd/entity/RunPod· home› entities› RunPod
grep -l @runpod /news/*.json | wc -l → 42

RunPod

mentions 42 type Organization page 1/3 feed RSS

// recent coverage 42 mentions

00:57
2026-09-29
brendanlong.com
ai-infrastructure

gpuc: A single-user GPU queue that doesn't require sudo

Independent developer Brendan Long released gpuc, an open-source single-user GPU queue that coordinates jobs across local, SSH-accessible remote, and rented RunPod GPUs without requiring sudo, publish…

19:38
2026-09-26
dev.to
ai-infrastructure

I changed nothing and my LLM server got 27% more expensive

A developer running an open-source CLI called Throttle against an unchanged local Ollama server (llama3.2:3b on a MacBook) measured the same configuration four times and saw output-token costs swing f…

10:21
2026-09-23
parity.io
ai-infrastructure

Self-hosting DeepSeek V4 for a software engineering org

Parity engineers ran a self-hosted inference trial of the open-weight DeepSeek V4 Flash model that handled more than 144,000 requests and almost 12.9 billion tokens from 25 engineers between 16 August…

15:32
2026-09-21
github.com
ai-infrastructure

Bootstrap scripts for ML workloads on a few GPU cloud providers

A GitHub repository published by tudormunteanu provides bootstrap scripts and a cheatsheet for setting up rented GPU cloud instances on RunPod, Vast, Nebius, and Enverge, aimed at making a rented GPU …

00:00
2026-09-11
mindstudio.ai
ai-infrastructure

RunPod Serverless Explained: Pay-Per-Second GPU API Deployment

RunPod Serverless lets developers deploy any Hugging Face model as an OpenAI-compatible API endpoint with per-second billing, scale-to-zero, and flash boot cold starts, according to RunPod. The platfo…

00:00
2026-09-10
mindstudio.ai
ai-products

Nex-N2.5 Mini Hands-On: Testing Next AGI's Agentic Model

Next AGI released Nex-N2.5 Mini, a multimodal agentic model built on a Qwen3.5 mixture-of-experts backbone, under an Apache 2.0 license on Hugging Face, requiring two 80GB-class GPUs to run. In hands-…

00:00
2026-09-10
mindstudio.ai
ai-products

How to Run Nex-N2.5 Mini Locally on RunPod (Dual H100 Setup)

Nex AGI's Nex-N2.5 Mini, the smaller model in the company's N2.5 agentic family, requires two 80GB-class GPUs such as H100s and consumed roughly 66GB of VRAM per card in testing on RunPod, according t…

13:26
2026-09-01
dev.to
artificial-intelligence

My Mac Is Useless for Local AI. My Windows Laptop Isn't.

A developer from Port Harcourt argues that local AI is not dead, citing his own experience where a Windows laptop with integrated graphics successfully ran a 7B coding model for an offline AI coding a…

05:40
2026-08-25
llmpanel.io
artificial-intelligence

LLMPanel Deploy vLLM to RunPod or Vast.ai Without Kubernetes

LLMPanel has launched an open-source platform that deploys large language models on any GPU cloud, including RunPod and Vast.ai, without requiring Kubernetes. The tool provisions containers, exposes O…

09:00
2026-08-24
infoworld.com
ai-infrastructure

Can neoclouds corner AI compute?

Neocloud revenue exceeded $25 billion in 2025, and Gartner predicts neoclouds will capture 20% of the $267 billion AI cloud market by 2030, according to Synergy Research Group and Gartner. These purpo…

03:19
2026-08-24
dev.to
generative-ai

A Rehearsal Is Only Cheap In Distribution

Scenematic's generative video pipeline uses cheap low-step sketches to select parameters, but out-of-distribution prompts can turn these surrogate scores into noise, leading to costly compounding fail…

17:01
2026-08-20
promptcube3.com
machine-learning

Colab free tier killed my 7B fine-tune — here's the autopsy

A developer's attempt to fine-tune Mistral-7B-v0.1 on Google Colab's free tier failed due to out-of-memory errors and a 2-hour session limit, with the runtime disconnecting at step 200 and a MemoryErr…

12:54
2026-08-18
discuss.huggingface.co
ai-infrastructure

Which GPU cloud provider are you actually using for inference?

A Hugging Face forum user is developing a tool to recommend GPU cloud providers based on specific use cases, citing frustration with comparing providers like RunPod, Vast.ai, and Lambda Labs. In respo…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics