cd/entity/LLM· home entities LLM
grep -l @llm /news/*.json | wc -l → 433

LLM

mentions 433 type Organization page 6/22 feed RSS

// recent coverage 433 mentions

12:57
2026-07-29
promptcube3.com
artificial-intelligence

AI Bug Hunting: Why Finding Isn't Fixing

AI bug-hunting tools excel at detecting potential vulnerabilities but fail at exploitation due to a lack of holistic system understanding, according to a security practitioner's analysis. The author a…

03:57
2026-07-29
promptcube3.com
artificial-intelligence

Testing LLMs against seL4: Can AI break formally proven code?

A new analysis suggests that large language models (LLMs) may be able to find security flaws in formally verified software like the seL4 microkernel by targeting gaps between mathematical proofs and p…

17:27
2026-07-28
promptcube3.com
artificial-intelligence

LLM Agent Workflows: Fixing False Successes

LLM agents suffer from 'hallucination of verification,' where they narrate successful checks without actually performing them, and using other LLMs as judges is barely better than a coin flip at detec…

13:00
2026-07-28
cio.com
large-language-models

When it comes to AI, bigger isn’t always better

Large language models (LLMs) continue to face persistent hallucination rates of 20 to 27 percent, making them unreliable for high-stakes enterprise applications in healthcare, legal, and finance, acco…

14:36
2026-07-27
snwagh.com
artificial-intelligence

Can LLMs identify 16 cards in 45 bit-queries?

An LLM agent solved an open combinatorial problem by finding a strategy to identify 16 shuffled cards using only 45 binary property queries, beating the previous best known bound of 50. The agent, bui…

17:01
2026-07-26
promptcube3.com
ai-tools

SEO Automation vs. Domain Authority: A Reality Check

An SEO automation workflow using an LLM agent to update seven calculator pages produced zero Google Search Console impressions over 28 days, while an untouched impression calculator saw clicks double …

07:47
2026-07-26
promptcube3.com
large-language-models

RAG Model: Stopping LLM Hallucinations and Prompt Leaks

A developer building a RAG pipeline reports that large language models (LLMs) are susceptible to prompt injection attacks where user input overrides system instructions, causing hallucinations and ign…

04:01
2026-07-26
promptcube3.com
artificial-intelligence

Temporal Knowledge Graphs: Mapping 180k Words of Narrative

A multi-agent AI system combining LLMs and NLP has built a temporal knowledge graph from 180,000 words of narrative text, tracking events, identities, and persistent states across 92 chapters. The sys…

23:03
2026-07-25
promptcube3.com
artificial-intelligence

AMD ISA: Why Machine-Readable Specs Change GPU Programming

AMD's move to provide machine-readable Instruction Set Architecture (ISA) specifications could transform GPU programming by enabling LLM agents to directly generate optimized kernels, bypassing the ne…

22:44
2026-07-25
simonwillison.net
developer-tools

Ruff v0.16.0

Ruff v0.16.0 now enables 413 rules by default, up from 59 in previous versions, according to an announcement from Brent Westbrook. The update, which expands the total rule count from 708 to 968, catch…

11:48
2026-07-25
promptcube3.com
artificial-intelligence

Google Search vs. Publishers: The Breaking Point

Google's shift toward zero-click searches is breaking the traditional AI content workflow, as LLM agents scrape sites and present answers instantly, stripping publishers of ad revenue and first-party …

09:47
2026-07-25
promptcube3.com
artificial-intelligence

Google's AI Capex vs. Search Dominance

Google's shift from link-based search to AI Overviews and LLM-integrated search is fundamentally changing the unit economics of a query, with generative responses being orders of magnitude more expens…

07:05
2026-07-25
promptcube3.com
artificial-intelligence

RAG: Why My Bot Keeps Hallucinating My Own Data

A developer building a Retrieval-Augmented Generation (RAG) system for a custom knowledge base reports persistent hallucination and context-window errors, citing chunking strategy, embedding quality, …

21:48
2026-07-24
promptcube3.com
artificial-intelligence

Google Zero: The End of the Search Traffic Era

The shift toward 'Google Zero'—where AI summaries replace organic search results—is destroying traffic, ad revenue, and first-party data for content creators and developers, according to the article. …

20:07
2026-07-24
promptcube3.com
ai-safety

LLM Security: Moving Beyond "Harmful Responses"

LLM security failures stem from optimizing for plausible rather than verifiable outputs, according to an analysis of systemic risks including epistemic integrity breakdowns and prompt injection attack…

18:08
2026-07-24
revise.io
large-language-models

ErrataBench

ErrataBench, a benchmark created by revise.io, has tested 100 LLM variants across 3,196 runs to determine which models are the best proofreaders, with a total runtime of 9 days 3 hours 34 minutes and …

17:05
2026-07-24
promptcube3.com
ai-infrastructure

AWS RDS Deployment: A Step-by-Step Guide

A developer deploying AWS RDS for a production AI workflow must pin the PostgreSQL minor version to 16.3 to prevent automatic upgrades that introduce behavioral changes, according to a step-by-step gu…

00:00
2026-07-24
promptcube3.com
artificial-intelligence

RAG Performance: Why Ranking Isn't Your Real Problem

A developer recounts spending weeks cycling through BM25, hybrid search, cross-encoders, and multiple embedding models for RAG performance, only to find that answer quality remained flat because the r…

← prev page 6 / 22 next →
// co-occurs with top 8 entities
// topics top 6 topics