cd/sources/runtimewire-auto-discovered· home sources Runtimewire (auto-discovered)
cat /sources/runtimewire-auto-discovered.feed | wc -l → 1228

Runtimewire (auto-discovered)

articles 1228 domain runtimewire.com → page 32/62 feed RSS
07:56
2026-07-23
runtimewire.com
ai-tools

VaultSort adds on-device AI scanner for sensitive Mac files

VaultSort founder Justin Haubrich shipped Guardian on July 8, an on-device AI scanner in VaultSort 4.4.0 that detects sensitive files such as passport images, tax records, and API keys on Macs without…

05:15
2026-07-23
runtimewire.com
large-language-models

Head to head: Google: Gemini 3.6 Flash vs Kimi K3

Kimi K3 defeated Google Gemini 3.6 Flash 115.0 to 106.0 in a head-to-head text task evaluation, with 94% confidence, according to a benchmark run by an unnamed tester using 12 fresh tasks scored by GP…

05:12
2026-07-23
runtimewire.com
large-language-models

Head to head: Google: Gemini 3.6 Flash vs GLM 5.2

Google's Gemini 3.6 Flash defeated GLM 5.2 by a score of 68.0 to 53.0 in a head-to-head comparison of instruction-following ability, winning 6 of 8 tasks with one tie and 96% confidence. The test, con…

05:11
2026-07-23
runtimewire.com
artificial-intelligence

Head to head: Google: Gemini 3.6 Flash vs DeepSeek-V4-Pro

Google's Gemini 3.6 Flash defeated DeepSeek-V4-Pro in a head-to-head text task evaluation, winning 5 tasks to 2 with 86% confidence and an overall score of 111.7 to 98.8, according to tests run by an …

05:10
2026-07-23
runtimewire.com
large-language-models

Head to head: Google: Gemini 3.6 Flash vs gpt-oss-120b

Google's Gemini 3.6 Flash defeated gpt-oss-120b 108.8 to 102.8 in a 12-task head-to-head benchmark, winning 5 tasks to 3 with 4 ties, according to scores from judge gpt-5.4. The win was driven by cons…

04:59
2026-07-23
runtimewire.com
artificial-intelligence

Head to head: Kimi K3 vs Anthropic: Claude Opus 4.8

Kimi K3 defeated Anthropic's Claude Opus 4.8 in a head-to-head Three.js coding challenge, scoring 6.5 to 2.1 on a single scored task. The test, judged twice by GPT-5.4 to cancel position bias, found K…

04:49
2026-07-23
runtimewire.com
generative-ai

Head to head: Heygen vs Gemini Omni Flash

Gemini Omni Flash beats Heygen 34.5 to 18.8 in a head-to-head video generation test, winning 3 of 4 tasks with 1 tie and 94% confidence, according to a comparison run by an unnamed tester using GPT-5.…

03:43
2026-07-23
runtimewire.com
artificial-intelligence

Substack brings Pangram AI scans to posts, Notes and comments

Substack CEO Chris Best launched a reader-triggered AI detector on July 21st, integrating technology from Brooklyn-based AI-detection startup Pangram to estimate how much of a post, note, comment, or …

02:57
2026-07-23
runtimewire.com
artificial-intelligence

Petals lists 405B Llama support for volunteer GPU inference

Petals, an open-source distributed inference system, now supports Llama 3.1 models with up to 405 billion parameters, allowing users to run large language models across a volunteer network of consumer…

02:09
2026-07-23
runtimewire.com
artificial-intelligence

CrucibleBench finds an LLM judge shifted rankings by six spots

Removing two classifier-dependent scoring dimensions from CrucibleBench shifted language model leaderboard positions by as many as six spots, researchers Benjamin Davis and Philip Mims reported. The e…

22:24
2026-07-22
runtimewire.com
large-language-models

Developers abandon Claude for cheaper, open Kimi‑K3

A Reddit user identifying as ComputeIQ announced on July 21st that they are dropping Anthropic's Claude for Kimi's K3 model, citing Claude's rising per-token cost, opaque usage limits on the Fable-5 t…

← prev page 32 / 62 next →