cd/sources/vincentschmalbach-auto-discovered· home› sources› Vincentschmalbach (auto-discovered)
cat /sources/vincentschmalbach-auto-discovered.feed | wc -l → 54

Vincentschmalbach (auto-discovered)

articles 54 domain vincentschmalbach.com → page 1/3 feed RSS
17:10
2026-09-30
vincentschmalbach.com
ai-agents

How to Make Videos with Claude Code

Developer Vincent Schumacher released explainroo, a free, MIT-licensed open-source framework that lets Claude Code and other AI coding agents such as Codex, Pi and OpenCode produce animated explainer …

06:33
2026-09-17
vincentschmalbach.com
ai-search

We Urgently Need a New Major Search Engine

Google's AI Overviews grew from 15% of searches to 43% in one year and AI Mode visits rose from 126 million in June 2025 to 279 million in May 2026, according to Similarweb data cited by Vincent Schma…

16:16
2026-09-16
vincentschmalbach.com
ai-tools

Setting Up Pi With DeepSeek V4.1 Flash on OpenRouter

A developer's audit of pi's session logs against OpenRouter's generation API found a supposedly cheap DeepSeek V4.1 Flash setup paying about three times what it should, because OpenRouter's default Au…

03:48
2026-09-16
vincentschmalbach.com
ai-infrastructure

Inference Is the Last LLM Moat

OpenAI and Anthropic's remaining competitive moat is subsidized inference, and losing it would cost them 50% or more of individual developer customers and small businesses, according to an analysis pu…

13:25
2026-09-15
vincentschmalbach.com
ai-agents

Built-In Tools vs. Custom Tools in LLM Agents

A technical comparison of built-in versus custom tools in LLM agents finds that provider-run tools such as OpenAI's and Anthropic's hosted web search execute the entire tool loop inside a single API r…

09:49
2026-09-13
vincentschmalbach.com
large-language-models

My First Week with GPT-6 Astra

Vincent Schmalbach returned to GPT-5.5 as his default Codex driver after one week with GPT-6 Astra, despite finding Astra more capable than GPT-5.6 Sol and on par with Fable on front-end work. Schmalb…

11:25
2026-08-11
vincentschmalbach.com
large-language-models

How Floating-Point Determinism Affects LLM Reproducibility

Floating-point arithmetic non-determinism can cause LLM inference to produce different outputs across runs, hardware, or provider routing, according to an analysis of numerical execution in large lang…

11:25
2026-08-11
vincentschmalbach.com
large-language-models

Can Provider Routing Change LLM Outputs?

Provider routing can change an LLM's output when a request reaches a different model version, fallback model, parameter configuration, precision level, inference engine, region, or runtime environment…

11:25
2026-08-11
vincentschmalbach.com
large-language-models

How Model Updates Break LLM Reproducibility

A hosted LLM's output can change when providers update models, safety controls, routing, or serving infrastructure without code changes, breaking reproducibility. Google's Gemini API documentation exp…

11:25
2026-08-11
vincentschmalbach.com
large-language-models

What Is a System Fingerprint in LLM APIs?

OpenAI's API documentation explains that the `system_fingerprint` field in LLM responses is a provider-generated marker for the backend configuration serving a request, not a reflection of the user's …

11:25
2026-08-11
vincentschmalbach.com
large-language-models

Can Tokenizer Changes Affect LLM Output?

Changing a tokenizer can alter an LLM's output, ranging from no visible difference to lower quality, different formatting, shorter usable context, or complete inference failure, according to Hugging F…

11:25
2026-08-11
vincentschmalbach.com
large-language-models

Is the Same Prompt Always the Same LLM Input?

The same visible prompt is not always the same input received by a large language model (LLM), according to an analysis of provider routing and request assembly. Identical visible prompts do not estab…

11:25
2026-08-11
vincentschmalbach.com
artificial-intelligence

How to Test a Nondeterministic LLM Application

Testing a nondeterministic LLM application requires treating it as a workflow with behavioral requirements across repeated runs, not as a function returning one exact string, according to a guide that…

09:13
2026-08-11
vincentschmalbach.com
large-language-models

What Is Batch Invariance in LLM Inference?

VLLM's batch-invariance documentation defines a feature ensuring a request produces the same inference result regardless of batch size, composition, request order, or scheduling under a fixed hardware…

09:13
2026-08-11
vincentschmalbach.com
large-language-models

Do Structured Outputs Make LLM Responses Deterministic?

OpenAI reported that its strict Structured Outputs feature achieved 100% schema-matching reliability in an internal evaluation for gpt-4o-2024-08-06, but the company and other providers such as Amazon…

09:13
2026-08-11
vincentschmalbach.com
large-language-models

How Does Mixture-of-Experts Routing Affect LLM Repeatability?

A new analysis from the Journal of Machine Learning Research explains that Mixture-of-Experts (MoE) routing can cause large language models to produce different outputs across runs due to discrete exp…

09:12
2026-08-11
vincentschmalbach.com
large-language-models

Does Setting a Seed Make LLM Output Reproducible?

Setting a seed does not guarantee identical large language model (LLM) output, according to an analysis of LLM inference. A seed initializes the pseudorandom number generator used in token sampling bu…

05:34
2026-08-11
vincentschmalbach.com
large-language-models

How to Roll Out a New LLM Model Version Safely

A new large language model (LLM) version can alter production behavior in unpredictable ways, so organizations should use a gated, progressive, reversible rollout that includes defining a production c…

page 1 / 3 next →