cd/entity/gpt-4o-mini· home› entities› gpt-4o-mini
grep -l @gpt-4o-mini /news/*.json | wc -l → 47

gpt-4o-mini

mentions 47 type Organization page 1/3 feed RSS

// recent coverage 47 mentions

00:11
2026-10-08
dev.to
ai-research

My Eval Said RAG Made Things Up. My Eval Was Wrong.

A developer maintaining django-explain-errors, a Django middleware that uses an LLM to explain unhandled exceptions, found that their RAG evaluation harness falsely flagged retrieval-augmented explana…

14:00
2026-10-06
arize.com
ai-agents

What agent traces can tell you without an LLM judge

Arize Phoenix contributor Ashwin Govind Ugale released tracelint, an open-source linter that runs deterministic structural checks on OpenInference traces and returns a CI-usable exit code, avoiding LL…

05:09
2026-10-05
dev.to
ai-tools

Screenshot Journal

A developer built Screenshot Journal, a mobile-responsive spatial canvas app that turns camera-roll clutter into an interactive scrapbook where users can drag, group, search, and annotate screenshots …

22:21
2026-09-26
dev.to
ai-tools

AI Powered Git Commit Assistant

A developer built cmt-cli, an open-source command-line tool that analyzes staged Git changes and generates commit messages using OpenAI models, defaulting to gpt-4o-mini. The tool, installable from Py…

00:00
2026-09-21
motherduck.com
ai-products

Introducing prompt_jev(): bringing Jev to Motherduck SQL

MotherDuck shipped prompt_jev(), an integration with TypeSafe AI's Jev model, which classified 100,000 AG News articles in 40 seconds at 89% accuracy for $0.50, versus $37.58 and 31 minutes 59 seconds…

21:38
2026-09-08
dev.to
artificial-intelligence

Date Slop: building a deliberately bad UX with AI

A developer built 'Date Slop', a deliberately frustrating AI-powered date-picker that parodies overzealous AI assistants. The project, which uses OpenAI's gpt-4o-mini, intercepts date input and forces…

07:50
2026-09-07
github.com
artificial-intelligence

Agent memory that forgets and distorts on purpose

SelMem, an open-source Rust project by jbsalles, introduces selective reconstructive memory for large language models, enabling two instances to diverge by forgetting, gilding, and anchoring a particu…

11:42
2026-09-04
sourcefeed.dev
artificial-intelligence

Evaluate and Debug RAG Pipelines with Ragas

Ragas 0.4.3 introduces an evaluation harness that scores RAG pipeline outputs on faithfulness, context precision, and answer relevancy, using LLM judges to identify whether failures originate from the…

14:19
2026-08-30
dev.to
developer-tools

I built CI for prompts, and the first bug was in the tests

A developer built Sentinel, a prompt regression gate for CI, during the Agent Harness Hackathon. Sentinel runs an eval suite against both versions of a changed prompt, accounts for run-to-run noise, a…

21:57
2026-08-27
promptcube3.com
artificial-intelligence

AI web scraping

AI web scraping uses large language models to interpret webpage layouts and extract data semantically, replacing brittle HTML selectors. The approach reduces maintenance but introduces token costs and…

04:00
2026-08-27
machinebrief.com
artificial-intelligence

Adaptive Triggering for Bias Correction in LLM Reasoning

A new arXiv preprint (2608.25379v1) proposes adaptive triggering for bias correction in LLM reasoning, framing intervention timing as an online change-point detection problem with CUSUM statistics. On…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics