cd/entity/Vercel AI Gateway· home entities Vercel AI Gateway
grep -l @vercel ai gateway /news/*.json | wc -l → 37

Vercel AI Gateway

mentions 37 type Organization page 1/2 feed RSS

// recent coverage 37 mentions

00:00
2026-08-21
digitalapplied.com
artificial-intelligence

DeepSeek V4 Flash Vision: Images, Same Price, Two Clocks

DeepSeek released deepseek-v4-flash-vision-exp, an experimental multimodal variant of its V4-Flash model, on August 21, 2026, matching text-only capabilities while accepting images at the same per-tok…

11:14
2026-08-20
getaivo.dev
developer-tools

Use your favorite coding agent with any model

Aivo released a command-line tool that lets developers launch coding agents such as Claude Code, Codex, and Pi with any model from providers including OpenRouter, Vercel AI Gateway, DeepSeek, Cloudfla…

06:15
2026-08-20
byteiota.com
developer-tools

Vercel fx: Open-Source Coding Agent Built in Zig Ships

Vercel Labs open-sourced fx, a coding agent CLI written in Zig that ships as a 7.8 MiB binary, cold-starts in 10 microseconds, and requires no runtime to install. The tool, currently at v0.0.4, is des…

08:14
2026-08-12
github.com
ai-tools

A minimal chatbot template built with Next.js

Vercel released a minimal chatbot template built with Next.js, the AI SDK, shadcn/ui, shadcn/react, shadcn/typeset, and the Vercel AI Gateway, featuring streaming chat with markdown rendering, tool ca…

00:00
2026-08-10
vercel.com
ai-tools

Simplified onboarding for deepsec

Vercel's open-source security review harness deepsec now supports one-command setup, allowing users to configure a repository and run its first security review with a single `npx deepsec init` command…

00:00
2026-08-06
vercel.com
generative-ai

Seedance 2.5 now available on Vercel AI Gateway

ByteDance's Seedance 2.5 video generation model is now available on Vercel's AI Gateway, enabling single-clip generation of up to 30 seconds with consistent camera movement and continuity, plus local …

00:00
2026-08-05
vercel.com
artificial-intelligence

Muse Spark 1.2 is now available on Vercel AI Gateway

Meta's Muse Spark 1.2 is now available on Vercel AI Gateway, featuring improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows. The model is desi…

00:00
2026-08-04
conductor.build
ai-tools

Conductor 0.79.0: Cloud Polish #1

Conductor 0.79.0, a cloud-focused update to the AI coding agent, introduces simplified onboarding, Vercel AI Gateway API key support, and an experimental Pierre editor for direct file editing in diffs…

00:00
2026-08-04
zackproser.com
artificial-intelligence

The Eval Harness: 60 Local and Cloud Coding Runs

A 60-run evaluation of coding agents found no quality winner among local DeepSeek V4 Flash, hosted DeepSeek via Vercel AI Gateway, and Claude Sonnet 5, with strict acceptance at 8/20, 9/20, and 5/20 r…

15:56
2026-08-03
runtimewire.com
artificial-intelligence

FLock adds DeepSeek V4 Flash to OpenAI-compatible API gateway

FLock has added DeepSeek V4 Flash to its OpenAI-compatible API Platform, offering developers access to a 284-billion-parameter mixture-of-experts model with a 1-million-token context window at discoun…

00:00
2026-08-02
zackproser.com
large-language-models

Sampling Inkling and the Alias Pattern

Thinking Machines' Inkling, a 975B total / 41B active MoE model with Apache 2.0 license and 1M context, is being trialed via a bash wrapper alias 'claude-inkling' that bridges Anthropic's CLI to the O…

00:00
2026-08-02
zackproser.com
artificial-intelligence

Practical Advice on Local Inference and Cloud LLMs - August 2026

A developer's two-week experiment with an M5 Max MacBook Pro with 128 GB of unified memory found that DeepSeek V4 Flash weights ran at 6 tokens per second under mainline llama.cpp but 30 to 40 tokens …

17:43
2026-07-22
dev.to
artificial-intelligence

OpenRouter vs Vercel vs LLMGateway Performance

A developer benchmarked time-to-first-token (TTFT) across AI gateways, finding LLM Gateway delivered the first content token 35% faster cold (906ms vs 1392ms) and 34% faster warm (814ms vs 1232ms) tha…

21:07
2026-07-20
runtimewire.com
ai-infrastructure

Ramp opens AI model router, says it cut internal LLM costs 30%

Ramp opened a closed beta for Ramp Router on July 20th, a gateway that directs AI requests across approved models and tracks costs, which the company says cut its internal large language model costs b…

09:20
2026-07-16
dev.to
large-language-models

Inkling MoE + Agent Safety: Token Efficiency Meets Reliability

Inkling, a decoder-only mixture-of-experts model with 1 trillion total parameters and 40 billion active per token, launched on Together Serverless, supporting native multimodal I/O and a reasoning_eff…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics