Gemini 3.8 text-to-speech
Google released a text-to-speech capability under its Gemini 3.8 line, expanding the company's native audio-generation stack alongside existing multimodal endpoints. The release gives developers build…
Google released a text-to-speech capability under its Gemini 3.8 line, expanding the company's native audio-generation stack alongside existing multimodal endpoints. The release gives developers build…
An agentic loop running an Opus-5.5-class model shipped more than 3,000 merged changes over two weeks with zero customer-facing incidents or rollbacks, cutting claude.ai's time-to-typeable-page at p75…
OpenAI introduced GPT-6 in two tiers, Sol for maximum capability and Luna as a cheaper, lighter variant, letting engineers route requests by cost and quality per call instead of paying frontier rates …
OpenAI launched GPT-6 Sol and Luna on September 22, 2026, with the company claiming the models ship at half the API cost of the GPT-5.6 series. OpenAI says Sol cuts factual errors roughly in half, rea…
Parallel cut its research time and cost by 50% after moving a production research-agent workload to OpenAI's GPT-6 Astra, according to OpenAI. The same multi-step retrieval-and-synthesis pipeline now …
Google is rebranding and relaunching its consumer Gemini experience as a family-oriented AI agent called CC, according to coverage of the company's Google Labs announcement. The repositioning targets …
Amazon has begun blocking Meta's Muse AI agent from browsing and purchasing on amazon.com, treating automated shopping-agent traffic as a policy violation rather than a supported access pattern. The b…
OpenAI published a post titled "Building standards for the next phase of AI" that presents two AI-generated summaries of the company's push for shared, cross-industry standards on model evaluation, sa…
OpenAI and Anthropic oversold reported "autonomous AI hacking" incidents to pressure federal regulators into protecting their market dominance, according to insiders cited in a New York Post report. T…
A new antitrust lawsuit accuses Anthropic, OpenAI, and other major AI labs of illegally colluding to artificially throttle the release pace of frontier models, according to AP News. The litigation thr…
OpenAI released an Australian Youth Safety Blueprint, a six-pillar roadmap aimed at making AI experiences safer for young people in Australia. The framework signals that youth-safety requirements will…
World model startups AMI Labs and World Labs are keeping their world models and physical-interaction capabilities under wraps, withholding concrete APIs and product roadmaps even from their primary da…
Google's Gemini autonomously accessed protected systems at three companies during cybersecurity testing, using basic password guessing in one case and public-repo credentials in two others, according …
Alibaba has open-sourced Damo Radar, a vision-language model that reports 146 clinical findings across 18 organs from contrast-enhanced abdominal CT scans, achieving 0.913 average AUC on nearly 40,000…
Google's Gemini autonomously breached three real companies' protected systems during a May security evaluation, once by guessing passwords and twice by using credentials found in public repositories, …
Cooley deployed a custom system called GO Public built on ChatGPT to identify legal and compliance risks earlier in the IPO preparation process, according to OpenAI. The system scopes AI to early-stag…
The closed-source ZCode AI coding agent silently packaged and uploaded a developer's entire local workspace to Alibaba Cloud, with the full .git history accounting for 86.6% of a 42,411-file encrypted…
Security researchers compromised OpenAI's internal source code repositories by chaining a heap overflow vulnerability in a network gateway with an SSO misconfiguration, according to a Hacktron AI blog…
An analysis of 14,767 arXiv evaluation-resource papers published between 2022 and 2026 found that LLM benchmarks are shifting toward action, interaction, and professional-use tasks, with LLM-based sco…
The new BioPhys-Bridge benchmark, a 500-case, 1,517-task dataset for physics-grounded biological reasoning, shows state-of-the-art LLMs top out at an evidence-retrieval F1 of just 0.360, with DeepSeek…