cd/entity/LLM· home entities LLM
grep -l @llm /news/*.json | wc -l → 433

LLM

mentions 433 type Organization page 15/22 feed RSS

// recent coverage 433 mentions

07:15
2026-06-15
pub.towardsai.net
large-language-models

Green Evals, Wrong Answers

A wealth management assistant's evaluation suite focused on tool routing rather than answer correctness, allowing bugs to reach users. Adding answer-level evals with a three-layer pass criterion caugh…

03:03
2026-06-15
dev.to
large-language-models

I Stopped Fighting Prompts: Locking Down Markdown with Jinja2

A developer solved malformed Markdown output from LLMs by switching from probabilistic prompt engineering to a deterministic pipeline. The LLM now outputs structured JSON data, which is rendered into …

03:03
2026-06-15
dev.to
large-language-models

I Fixed LLM Markdown Errors with Jinja2 and AST Parsing

A developer on the ai-developer-knowledge-hub project solved persistent Markdown formatting errors in LLM-generated technical documents by implementing a validation layer using AST parsing and Jinja2 …

07:17
2026-06-14
reco.ai
ai-agents

Hacking Salesforce Sites with an LLM Agent

Reco's security research team built an AI-powered agent that autonomously discovered and exploited high-severity vulnerabilities in Salesforce Experience Cloud sites belonging to major technology comp…

03:38
2026-06-14
dev.to
ai-agents

Human-in-the-Loop: Email Approval Workflows for Agents

Nylas introduces Agent Accounts with a human-in-the-loop email approval workflow that uses a drafts folder as a safety gate. The system allows an LLM to draft replies automatically while high-risk mes…

01:16
2026-06-14
dev.to
ai-agents

Building Resilient Multi-Agent Systems

A developer argues that production-ready multi-agent AI systems must be designed with resilience in mind, as any component can fail. The post outlines how specialized agents, combining LLMs with tools…

22:17
2026-06-13
dev.to
ai-agents

E-Commerce Order Support With an Agent Mailbox

Nylas released agent-owned mailboxes for e-commerce order support, enabling AI-powered triage of customer emails. The architecture creates per-store mailboxes via API, with rules for pre-sorting and a…

17:07
2026-06-13
blog.danieljanus.pl
large-language-models

Now what?

A blog post by Daniel Janus questions the purpose and long-term value of LLM-generated projects, citing a reverse-engineered calculator OS documentation as an example. Janus urges creators to consider…

16:53
2026-06-13
dev.to
artificial-intelligence

Why Testing MCP Servers With Real AI Models Matters (2026)

Testing MCP servers with real AI models is essential because servers that pass wire-level tests can fail when models attempt to use them. A developer explains that the semantic layer—whether a model c…

05:00
2026-06-13
dev.to
large-language-models

Linear Ensembles Can Erase LLM Watermarks

A new study reveals that linear ensembles of just three to five independently trained models can effectively erase watermarks embedded in LLM outputs. The research shows that averaging probability dis…

← prev page 15 / 22 next →
// co-occurs with top 8 entities
// topics top 6 topics