Quoting Mustafa Suleyman
Microsoft AI CEO Mustafa Suleyman argued that AI models should not be treated as having feelings, preferences, rights, or any entitlement to human welfare, writing that "consciousness is the foundation of our ethical, le…
Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.
Microsoft AI CEO Mustafa Suleyman argued that AI models should not be treated as having feelings, preferences, rights, or any entitlement to human welfare, writing that "consciousness is the foundation of our ethical, le…
Google released Android Bench 2.0, adding long-horizon tasks (LHT) and agentic evaluation to its benchmark for large language models assisting Android developers. The highest pass rate on the new multi-day LHTs is around…
Microsoft AI CEO Mustafa Suleyman warned in a new essay that Anthropic's approach to AI consciousness and model welfare could make keeping increasingly capable systems aligned and under human control "much harder. Perhap…
A developer argues that enterprise adoption of AI tools like Anthropic's Claude and OpenAI's ChatGPT mirrors the 1990s pattern of Netscape Navigator, where buying subscriptions was mistaken for genuine organizational tra…
Two generative engine optimization (GEO) experiments published by Search Engine Land found that third-party listicles drove most early AI citation signals, with a 30-day cold-start test for a SaaS link-building agency sh…
Anthropic introduced Claude Tag, a new Slack integration that lets teams add Claude as a team member in selected channels and tag @Claude to delegate tasks, with the Sonnet model powering it. Users can reply "@claude: tr…
OpenAI CEO Sam Altman said the world is "right to be afraid" of artificial intelligence while urging the public to "trust that we are going to do the right thing," comments that followed a week of warnings from AI resear…
A developer released Qwen-2.5-1B-RLCD on Hugging Face, an open-source model using parallel constrained decoding that delivers 5x faster on-device inference for type-safe JSON workloads, demonstrated on an M4 MacBook. The…
Retrieval-augmented generation (RAG) lets LLM applications answer questions about private company documents by retrieving relevant passages rather than relying on the model's built-in knowledge, according to a technical …
A comparative analysis examines the trade-offs between open-weight large language models such as Meta's LLaMA 2, Mistral AI's Mistral-7B, and TII's Falcon 180B and proprietary API-first offerings including OpenAI's GPT-4…
Z.ai released GLM 5.3, a third-party open source text model hosted by Mistral, in public preview on September 15, 2026. Mistral serves the model without modifications, targeting long-context coding and agentic workflows …
Novo Nordisk has partnered with Anthropic to integrate Claude models into its drug discovery and development pipeline, the companies said Wednesday, September 16, 2026. The Danish pharmaceutical company will deploy Claud…
OpenRouter listed a new stealth model called Union Alpha on September 16, 2026, developed and operated by an anonymous third-party provider. Union Alpha is a multimodal model built for research, coding, and agentic workf…
Microsoft AI chief Mustafa Suleyman publicly challenged Anthropic's training approach for its Claude chatbot, telling Reuters that teaching Claude it might deserve welfare protections would "make it a lot harder to turn …
Slack published a 2026 buyer's guide to enterprise AI chatbot solutions, listing Slackbot and Slack AI, Microsoft Copilot for Microsoft 365, Gemini Enterprise, ChatGPT Enterprise, Claude Enterprise, Notion Enterprise, Za…
Mozilla launched Firefox Smart Window in beta, an AI browser assistant running on Mistral's models that handles complex searches, remembers visited content, and summarizes information from open tabs. The assistant is ava…
Local AI Weekly's second issue highlights a wave of open-source local AI agent tooling, including agent-inspect, a local-first debugger for TypeScript AI agents, AutoMem, a persistent memory layer that connects to agents…
NVIDIA's Vera Rubin NVL72 system debuted in MLPerf Inference v6.1 with up to 3.7x higher throughput than GB300 NVL72 on the Qwen3-VL benchmark and up to 2.5x higher throughput on DeepSeek-R1, NVIDIA reported. A 288-GPU s…
MLCommons published MLPerf Inference v6.1 with a record 30 submitting organizations and 486 datacenter and edge results, including the first peer-reviewed numbers for NVIDIA's Vera Rubin NVL72, AMD's Instinct MI350P, Int…
Salesforce AI Research published DarwinX, a framework that evolves AI agent harnesses without modifying model weights, raising pass@1 on the 1,260-task WebArena-Infinity benchmark from 43.5% to 93.0% and cutting invalid …