# AI agents can now make real biology finds from one broad prompt

> Source: <https://www.vibeleaderboard.ai/intel/brief/2026-09-24>
> Published: 2026-09-24 11:09:15+00:00

Anthropic's life sciences lab gave Claude only a high-level prompt, and it found a novel enzyme system with CRISPR-like repeated DNA sequences in a jumbo phage. CRISPR pioneer Feng Zhang called the result genuinely intriguing.
Read: Anthropic's life sciences lab gave Claude only a high-level prompt, and it found a novel enzyme system with CRISPR-like repeated DNA sequences in a jumbo phage. CRISPR pioneer Feng Zhang called the result genuinely intriguing.
Read: NVIDIA's SWE-Serve benchmark gives coding agents 53 real SGLang inference-engineering tasks. Patches that pass conventional tests fail live-serving checks about a third of the time, and pass@1 across 11 models ranges from 35% to 75.5%.
Read: Google DeepMind released Gemini 3.8 Flash and Flash-Lite text-to-speech models that build custom voices from prompts and take line-by-line direction on emotion and pacing, across AI Studio, the Gemini API and Vercel's AI Gateway.
Read: At Meta Connect 2026, Muse moved to voice and real-time video and passed ChatGPT in the App Store. Meta also showed new AI glasses and explained Private Processing, which runs heavier models in attested cloud VMs.
Read: New testing shows Claude Code 2.1.277 and later load AGENTS.md only behind a remote feature flag, so turning off telemetry silently skips the file. A one-line CLAUDE.md containing @AGENTS.md restores it.
Read: DrivingBench gave frontier models control of a real Toyota Corolla's steering, throttle and brakes on a cone course. GPT-6 Astra reached 100% progress while Claude Fable 5.1 and Grok 4.6 stalled well short.
Read: OpenRouter's analysis finds Moonshot's Kimi K3 license lacks OSI approval and adds a Model-as-a-Service revenue gate and a UI attribution requirement above certain scale, despite broad use and resale rights.
