cd/entity/Whisper· home entities Whisper
grep -l @whisper /news/*.json | wc -l → 150

Whisper

mentions 150 type Organization page 4/8 feed RSS

// recent coverage 150 mentions

17:04
2026-07-13
sourcefeed.dev
artificial-intelligence

SpeechAnalyzer Beats Whisper On Apple Silicon

Apple's new SpeechAnalyzer engine, introduced with iOS 26 and macOS 26, beats OpenAI's Whisper Small on clean English read speech with a 2.12% word error rate on LibriSpeech test-clean versus Whisper …

00:00
2026-07-12
martino.im
ai-tools

Summarize

Martin Piaggi built Summarize, an open-source tool that converts videos from YouTube, TikTok, Instagram, X, Reddit, Facebook, and local files into structured notes using a pipeline of transcription (Y…

17:09
2026-07-11
byteiota.com
artificial-intelligence

GPT-Live-1: OpenAI’s Full-Duplex Voice Model, Explained

OpenAI launched GPT-Live-1 on July 8, a full-duplex voice model that listens and speaks simultaneously, using a two-layer architecture with a continuous interaction layer and a delegation layer that h…

15:02
2026-07-11
sourcefeed.dev
large-language-models

The Token-Saving Architecture of LLM Video Ingestion

The open-source claude-video tool uses client-side orchestration with yt-dlp, ffmpeg, and Whisper to make multimodal video analysis practical for LLMs by optimizing frame extraction and token usage. I…

03:54
2026-07-11
machinebrief.com
natural-language-processing

Luxembourgish AI: Giving Voice to a Small Language

Researchers are using text-to-speech systems to create synthetic training data for spoken question answering in Luxembourgish, a low-resource language. By translating existing QA resources and synthes…

16:09
2026-07-10
machinebrief.com
artificial-intelligence

SAMPA: A Leap Forward for Brazilian Portuguese Prosody

Researchers introduced SAMPA, a Whisper-based model for prosodic segmentation in Brazilian Portuguese, achieving F1 scores of 0.731 on held-out test data and 0.796 on the MuPe-Diversidades dataset. Th…

04:00
2026-07-10
machinebrief.com
artificial-intelligence

Best-of-$N$ TTS Evaluation is Confounded by ASR Family Alignment

Researchers found that best-of-N text-to-speech evaluation is confounded by alignment between ASR verifier and evaluator families, with same-family pairs recovering 2-3× more oracle headroom than cros…

05:31
2026-07-07
kamalg2.substack.com
artificial-intelligence

AI STS Stack for underserved languages paper

A developer evaluated real-time voice AI stacks for underserved languages, comparing latency and accuracy of various speech-to-text, text-to-speech, and translation models. The analysis found that a c…

← prev page 4 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics