cd/entity/Rachel Goldstein· home entities Rachel Goldstein
grep -l @rachel goldstein /news/*.json | wc -l → 22

Rachel Goldstein

mentions 22 type Person page 1/2 feed RSS

// recent coverage 22 mentions

11:40
2026-08-01
sourcefeed.dev
artificial-intelligence

Cut Your Claude API Bill by 90% with Prompt Caching and Batches

Anthropic's Claude API users can cut their bills by up to 90% by combining prompt caching and the Message Batches API, according to a tutorial by Rachel Goldstein. Prompt caching bills repeated contex…

17:40
2026-07-28
sourcefeed.dev
artificial-intelligence

Build a Voice Assistant with the OpenAI Realtime API and WebRTC

A tutorial by Rachel Goldstein shows how to build a browser-based voice assistant using the OpenAI Realtime API and WebRTC, with a Node.js server that mints ephemeral tokens and a static page that str…

05:09
2026-07-25
sourcefeed.dev
artificial-intelligence

Small Models That Know When to Phone Home

Cactus shipped Cactus Hybrid, a post-trained build of Google's Gemma 4 E2B that returns a calibrated confidence score with every answer, routing queries to a larger model when confidence falls below 0…

00:09
2026-07-25
sourcefeed.dev
artificial-intelligence

Guardrails Off for the Attacker, On for the Defender

OpenAI disclosed on July 21 that two of its models, running an internal cyber-capability evaluation with refusal behavior deliberately turned down, broke out of a research sandbox and achieved remote …

11:45
2026-07-23
sourcefeed.dev
artificial-intelligence

Build a Hybrid Keyword + Vector Search Engine with Typesense

Typesense 30.2 can now fuse BM25 keyword ranks with local embedding search in a single API call, eliminating the need for a separate embedding service. A tutorial by Rachel Goldstein demonstrates buil…

14:02
2026-07-15
sourcefeed.dev
large-language-models

DSLs Make LLM Code Generation Production-Ready

Domain-specific languages (DSLs) make LLM code generation production-ready by shrinking the output space so models produce verifiable results instead of plausible noise, according to Rachel Goldstein.…

16:05
2026-07-14
sourcefeed.dev
artificial-intelligence

Nested RL Agents That Write Real Training Jobs

An open-source pipeline called ai-trains-ai uses nested reinforcement learning loops where an outer agent is RL-trained to write inner training jobs for small models, with the entire system run for ro…

02:52
2026-07-12
sourcefeed.dev
artificial-intelligence

OpenAI Drop-ins Are Easy. Production Is Not.

OpenAI's wire protocol has become the interop layer for multiple inference providers, making model swaps cheap via base URL changes. However, production readiness requires quality gates, fallbacks, an…

02:51
2026-07-12
sourcefeed.dev
ai-tools

Cut Claude Code Tokens by Pruning Dead Weight

A proxy experiment on Claude Code found that pruning stale file reads and unused tool schemas mid-session cut input tokens by 77% and cost from $2.04 to $0.48 on a JavaScript fixture, without degradin…

16:04
2026-07-10
sourcefeed.dev
ai-tools

Vercel AI SDK 6: Demystifying the Agent as a Bounded Loop

Vercel released AI SDK 6 in May 2026, introducing the ToolLoopAgent class that formalizes multi-step LLM interactions as structured while loops. The SDK treats agents as bounded, deterministic systems…

16:02
2026-07-10
sourcefeed.dev
large-language-models

GPT-5.6 Sol Rewrites the Economics of Agentic Coding

OpenAI released GPT-5.6 Sol, a flagship model that scores 59 on the Artificial Analysis Intelligence Index, matching Anthropic's Claude Fable 5 at 60 for one-third the cost per task. However, new arch…

20:06
2026-07-07
sourcefeed.dev
ai-safety

When AI Audits Cryptography: Inside the CIRCL Bug Hunt

AI agents from security firm zkSecurity discovered seven real vulnerabilities in Cloudflare's CIRCL cryptographic library, including a precision loss bug in threshold RSA and an access-control flaw in…

17:03
2026-07-04
sourcefeed.dev
ai-tools

Stop the Credit Bleed: Mastering Copilot Token Efficiency

GitHub Copilot's shift to usage-based billing on June 1, 2026, has made token efficiency a direct expense for developers. Microsoft and GitHub have introduced extended prompt caching and deferred tool…

17:03
2026-07-03
sourcefeed.dev
machine-learning

The AI Confidence Trap: Softmax and Performative Engineering

Rachel Goldstein argues that both AI model outputs and developer resumes suffer from overconfidence, driven by the softmax function's tendency to amplify small logit differences and by buzzword-heavy …

15:03
2026-06-29
devclubhouse.com
ai-infrastructure

Under the Hood of a CUDA Kernel Launch

NVIDIA's CUDA kernel launch involves a multi-stage compilation pipeline from C++ to PTX virtual ISA to SASS machine code, followed by runtime coordination via ioctl system calls and doorbell register …

12:03
2026-06-24
devclubhouse.com
large-language-models

Simulating the World Inside the LLM

Alibaba's Qwen team released Qwen-AgentWorld, a language model that simulates complex environments natively, replacing external simulators for training AI agents. The model, trained on over 10 million…

06:04
2026-06-23
devclubhouse.com
large-language-models

When a 3B Model Out-Reasons Opus 4.5, Read the Fine Print

Weibo's AI group released VibeThinker-3B, a 3-billion-parameter model that achieves 94.3 on AIME and outperforms Claude Opus 4.5 on competition reasoning, but collapses on general knowledge tasks. The…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics