cd /news/large-language-models/how-much-cost-would-it-take-to-run-a… · home topics large-language-models article
[ARTICLE · art-68172] src=news.ycombinator.com ↗ pub= topic=large-language-models verified=true sentiment=· neutral

How much cost would it take to run a LLMs locally

Running large language models like Qwen or GPT locally requires a dedicated machine with sufficient RAM/VRAM, and costs vary based on hardware; smaller variants of Gemma or Qwen Coder are recommended for local use, while hosted inference providers offer faster performance at a small cost.

read1 min views1 publishedJul 22, 2026

If I want to run LLMs model like Qwen or gpt locally how much would it cost me and I want to connect it to my main website creating an API link, also which models would be best First, check what LLMs your system can actually handle based on your RAM/VRAM, then choose the model variant accordingly.

If you want to use the model through an API for your website, a good setup would be a dedicated machine/server to host it, since running LLMs locally can consume a lot of memory. Depending on your hardware, you can look at smaller variants of Gemma or Qwen Coder.

Another option is to use a hosted inference provider. It’ll cost a little, but you’ll usually get much faster inference compared to running locally especially if your system isn’t high-end.

── more in #large-language-models 4 stories · sorted by recency
── more on @qwen 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/how-much-cost-would-…] indexed:0 read:1min 2026-07-22 ·