cd /news/ai-infrastructure/show-hn-fast-inference-for-deep-seek… · home topics ai-infrastructure article
[ARTICLE · art-136406] src=coralbricks.ai ↗ pub= topic=ai-infrastructure verified=true sentiment=↑ positive

Show HN: Fast inference for deep seek flash v4.1 469 tok/s for coding

Coral Bricks, a Seattle-based inference platform founded on 2026-06-12 and backed by Afore Capital and Foundations Accelerator, launched with a claim of 469 tokens per second for DeepSeek Flash v4.1 on coding workloads. The company serves open models including Kimi, GLM and gpt-oss behind an OpenAI-compatible API, offering near-zero rate limits and free cached input tokens for long-running coding and research agents. Coral Bricks is led by CEO Hitesh Jain and Head of Engineering Divy Vasal.

read1 min views1 publishedSep 21, 2026
Show HN: Fast inference for deep seek flash v4.1 469 tok/s for coding
Image: source

Coral Bricks is an inference platform for coding and research agents that plan, call tools and reason over large contexts. It serves open models such as Kimi, GLM and gpt-oss behind an OpenAI-compatible API, with multiple times the tokens per second of a typical provider, near-zero rate limits, and cached input tokens free — so long-running agent workloads finish on time instead of queuing.

Company #

- Founded: 2026-06-12
- Headquarters: Seattle, WA, US

Founders #

  • Hitesh Jain — CEO
  • Divy Vasal — Head of Engineering
── more in #ai-infrastructure 4 stories · sorted by recency
── more on @coral bricks 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/show-hn-fast-inferen…] indexed:0 read:1min 2026-09-21 ·