cd/entity/LangSmith· home entities LangSmith
grep -l @langsmith /news/*.json | wc -l → 82

LangSmith

mentions 82 type Organization page 2/5 feed RSS

// recent coverage 82 mentions

15:47
2026-07-28
promptcube3.com
artificial-intelligence

Why my RAG pipeline kept hallucinating outdated API docs

A developer building a RAG-powered documentation bot for a TypeScript framework found that the bot kept hallucinating outdated API docs because vector search retrieved high-similarity chunks from old …

13:38
2026-07-28
mutagent.io
ai-agents

AI Engineers' Favourite AI Engineer

Mutagent, an AI engineer that reads traces, diagnoses failures, and ships fixes on a loop, launched today in research preview as a free tool. The company reports that running it on an AI recruiting co…

00:00
2026-07-28
mutagent.io
ai-agents

We're your AI Engineer's favorite AI Engineer

Mutagent, an AI engineer that reads traces, finds failures, and ships fixes on a loop, is now live in research preview and free to try. In production tests, it reduced an AI recruiting company's month…

19:08
2026-07-27
openvisor.vercel.app
developer-tools

Show HN: Openvisor – A Session Explorer for OpenCode

A developer released Openvisor, a local-first session explorer for OpenCode that lets users inspect tool calls, subagent use, and other session details from exported JSON files stored in the browser's…

04:50
2026-07-25
promptcube3.com
artificial-intelligence

AI ROI: Why Enterprises are Pivoting from Hype to Utility

Enterprises are pivoting from AI hype to utility as CFOs demand measurable ROI, with many early proof-of-concept projects failing to meet KPIs on cost per resolution or hours saved per employee, accor…

04:04
2026-07-25
promptcube3.com
large-language-models

Model Swapping: Why "Vibe Checks" Fail for LLM Agents

A new protocol for validating large language model (LLM) agent model swaps uses a 20-minute diffing process to catch regressions before production, according to a developer who built the open-source t…

00:00
2026-07-24
chaliy.name
ai-tools

You Do Not Need a Server for Evals

Evals for AI projects like coding agents and sandboxed bash do not require a dedicated server or platform, according to developer Everruns. Datasets, runners, and results can be stored and versioned d…

09:31
2026-07-12
startupfortune.com
ai-agents

How to Evaluate AI Agents Before You Ship Them to Real Users

Most founders shipping AI agents lack systematic evaluation methods, leading to public failures like Chevrolet's chatbot agreeing to sell a car for $1 and McDonald's AI drive-thru adding bacon to ice …

← prev page 2 / 5 next →
// co-occurs with top 8 entities
// topics top 6 topics