cd/entity/Promptfoo· home entities Promptfoo
grep -l @promptfoo /news/*.json | wc -l → 15

Promptfoo

mentions 15 type Organization feed RSS

// recent coverage 15 mentions

12:53
2026-08-16
promptcube3.com
artificial-intelligence

AI red teaming tools

Developers often rely on manual, ad-hoc testing for LLM features, but a more systematic approach using automated red teaming tools like Giskard and Promptfoo can uncover critical vulnerabilities quick…

20:59
2026-08-04
dev.to
large-language-models

How EvalPort's Grader System Works: 11 Types for LLM Evaluation

EvalPort introduces a grader system with 11 types for LLM evaluation, including exact_match, semantic_similarity, llm_judge, and custom, designed to be framework-agnostic and self-describing. The syst…

03:40
2026-07-30
dev.to
large-language-models

OpenEval: Why LLM Evaluation Needs a Standard Format

OpenEval, a new open-source project, aims to standardize LLM evaluation by defining a portable JSON Schema for test cases, graders, and results. The project provides SDKs, a CLI, and converters for po…

09:00
2026-07-29
pydantic.dev
ai-agents

The best AI agent optimization platforms in 2026

Pydantic Logfire ranks as the best AI agent optimization platform in 2026, according to a Pydantic analysis that evaluated tools on their ability to close the loop by proposing and shipping production…

21:56
2026-07-23
promptcube3.com
ai-safety

Open Source AI Community, AI red teaming tools, Wi

AI red teaming tools like Giskard, DeepEval, and Promptfoo automate adversarial testing to systematically find edge-case failures in model logic, moving beyond manual 'vibe checks' that risk PR disast…

06:30
2026-07-14
agentic.tracebit.com
ai-safety

Context bombs: stopping AI attackers in their tracks

A new defensive technique called a 'context bomb' — a short string hidden in decoy resources that triggers safety guardrails in offensive AI agents — reduced autonomous cyberattack success by roughly …

10:13
2026-06-25
dev.to
large-language-models

Evaluating a C# LLM Eventparser with Promptfoo

A developer built a C# LLM event parser called EventParser and tested it using Promptfoo's LLM-as-a-judge evaluation. The project separates the prompt file from the code, allowing Promptfoo to test th…

00:00
2026-06-12
runagentrun.co.uk
artificial-intelligence

OpenAI acquires Ona to make Codex persistent

OpenAI acquired cloud startup Ona on 11 June 2026 to integrate its secure, persistent agent workspace into Codex, the company's coding assistant. The deal aims to support Codex's growing enterprise us…

04:17
2026-05-20
dev.to
large-language-models

Promptfoo: LLM Red Teaming Against OWASP Top 10

The open-source tool Promptfoo, acquired by OpenAI in March 2026, maps its 155 attack plugins to the OWASP LLM Top 10 2025 list for structured red teaming of LLM-powered products. It details the 2025 …

02:54
2026-05-20
dev.to
large-language-models

Prompt Versioning and Prompt Management for Engineering Teams

As engineering teams iterate on prompts for large language models, they face challenges similar to managing configuration files, requiring dedicated storage, versioning, sharing, and security. It revi…

// co-occurs with top 8 entities
// topics top 6 topics