cd/entity/DeepEval· home› entities› DeepEval
grep -l @deepeval /news/*.json | wc -l → 32

DeepEval

mentions 32 type Organization page 2/2 feed RSS

// recent coverage 32 mentions

21:56
2026-07-23
promptcube3.com
ai-safety

Open Source AI Community, AI red teaming tools, Wi

AI red teaming tools like Giskard, DeepEval, and Promptfoo automate adversarial testing to systematically find edge-case failures in model logic, moving beyond manual 'vibe checks' that risk PR disast…

19:27
2026-07-22
dev.to
artificial-intelligence

An LLM judge is a biased instrument, not a measurement

A developer found that an LLM judge gave opposite results for the same eval run on consecutive days due to position bias, one of three systematic biases documented in the 2023 paper "Judging LLM-as-a-…

15:30
2026-07-11
dev.to
large-language-models

How to Add Evals to an LLM Feature

A developer explains how to add evals to an LLM feature, using an outbound AI calling agent as an example. The process involves defining a business outcome metric, curating a representative dataset of…

18:40
2026-06-11
magnus919.com
artificial-intelligence

AI Evals 101: Stop the Slop

AI evaluations are essential for distinguishing working systems from broken ones, yet many companies ship AI into production without systematic quality checks, relying on 'vibe checks' instead. A four…

21:49
2026-05-26
letsdatascience.com
artificial-intelligence

Jenkins Continues Development of AI Chatbot for Resources

Mallikarjun G D and Daniele Caldarigi published Jenkins blog posts on May 26, 2026, detailing two GSoC 2026 projects extending the Jenkins ecosystem with AI chatbot plugins. G D's plugin adds an LLM-a…

← prev page 2 / 2
// co-occurs with top 8 entities
// topics top 6 topics