cd/sources/dev-to· home sources Dev.to
cat /sources/dev-to.feed | wc -l → 18708

Dev.to

articles 18708 domain dev.to → page 268/936 feed RSS
23:01
2026-08-14
dev.to
ai-infrastructure

Why Making AI Answer Faster Is Worth $1.5 Billion

Fireworks AI raised $1.5 billion at a $17.5 billion valuation, signaling a shift in AI investment toward inference optimization. The company, which serves over 40 trillion tokens daily and generates $…

22:40
2026-08-14
dev.to
ai-agents

How to Review AI-Generated Code with Multiple AI Agents

Entire, a developer tooling company, introduced a workflow that lets developers review AI-generated code with multiple AI agents from different model families, such as using Codex to build and Claude …

22:37
2026-08-14
dev.to
ai-agents

Agent Memory, Part 3: Perfection Is a Unicorn

In the final part of a series on agent memory, developer John Onlee describes rebuilding his coding agent Monet and the harness to maximize the model's own abilities, and argues that the key to agent …

22:17
2026-08-14
dev.to
developer-tools

How to Enable Entire

Entire, a tool that captures the context behind AI-assisted code changes and links it to Git history, has introduced a new setup process. Developers can enable Entire by running 'entire enable -y' fro…

21:35
2026-08-14
dev.to
large-language-models

Your memory layer is lying to you (and your LLM agrees)

An engineer's verify-on-read experiment with live LLMs found that cheap flash-tier models like qwen3.6-flash and qwen3.7-flash achieve zero false-accept rates on memory contamination checks at a fract…

21:16
2026-08-14
dev.to
artificial-intelligence

Google lowers Gemini 3.7 Flash costs for developers

Google has launched Gemini 3.7 Flash, a developer-focused AI model with reduced production pricing, now at $0.75 per million input tokens and $3.75 per million output tokens, roughly half the cost of …

20:45
2026-08-14
dev.to
artificial-intelligence

Let a Free Model Try to Break Your API Before Your Users Do

MonkeyCode's product outreach demonstrates a technique for using a free language model to generate adversarial API test payloads. The approach, detailed in a script, asks the model to produce hostile …

20:43
2026-08-14
dev.to
developer-tools

Serving Gemma4 with Rust on vLLM 🦀

A developer detailed how to build and run vLLM's Rust frontend on an AWS EC2 G5g instance with Graviton2 and an NVIDIA T4G GPU. The tutorial highlights that vLLM now requires a Rust toolchain for sour…

← prev page 268 / 936 next →