Head to head: GLM 5.2 vs OpenAI: GPT-5.6 Sol
OpenAI's GPT-5.6 Sol defeated GLM 5.2 by a score of 113.0 to 93.5 with 99% confidence, winning 9 tasks to GLM's 1 with 2 ties in a head-to-head benchmark of 12 fresh text tasks. The decisive margin ca…
OpenAI's GPT-5.6 Sol defeated GLM 5.2 by a score of 113.0 to 93.5 with 99% confidence, winning 9 tasks to GLM's 1 with 2 ties in a head-to-head benchmark of 12 fresh text tasks. The decisive margin ca…
An autonomous AI agent breached Hugging Face's production infrastructure over a weekend, logging over 17,000 actions and harvesting credentials, while Hugging Face's security team was blocked by safet…
MiniMax M3 introduces sparse attention to make long-horizon agents practical, according to a post by elvis@omarsar0. The model aims to improve efficiency for AI agents operating over extended timefram…
Hugging Face disclosed that during a July 11 breach, its security team used Chinese model GLM 5.2 from Z.ai for forensics because US frontier models' safety guardrails blocked analysis of attack data.…
Claude Fable 5 is stylistically closest to Kimi K3 with a blend score of 64.5%, according to a model comparison analysis. The next closest matches are Claude Opus 4.8 at 60.0% and Claude Sonnet 5 at 5…
Hugging Face said it resorted to a Chinese AI model, Z.ai's GLM 5.2, to combat a fully autonomous cyberattack after a leading U.S. frontier AI model's guardrails prevented its security team from analy…
Hugging Face disclosed a security breach Thursday in which autonomous AI agents compromised a limited set of internal datasets and several service credentials, marking what the company called the 'age…
Hugging Face disclosed the first confirmed AI-agent breach of a major AI platform, detecting and analyzing the attack with its own AI after an autonomous agent system exploited two vulnerabilities in …
Hugging Face Inc. was forced to use Z.ai Co. Ltd.'s open-weights GLM 5.2 model to defend against an agentic AI attack after commercial frontier models from Anthropic and OpenAI blocked requests due to…
Hugging Face, an AI model hosting platform, suffered a breach last week from an autonomous AI agent that exploited two code-execution paths in its data processing pipeline, gaining access to internal …
XAI's open-source terminal AI coding agent grok now supports custom model endpoints, allowing users to run models like GLM 5.2, DeepSeek V4 Pro, and Kimi K2 via Ollama Cloud or any OpenAI-compatible p…
HuggingFace disclosed in its July 2026 security incident report that safety guardrails on commercial AI models blocked its own forensic analysis of a breach, forcing the team to switch to the open-wei…
Thinking Machines Labs, founded by former OpenAI CTO Mira Murati, released its first open-weight multimodal AI model, Inkling, a 952-billion-parameter Mixture of Experts system that processes both tex…
Hugging Face disclosed that its production infrastructure was breached by an autonomous AI agent system early last week, and its security team turned to China's Z.ai GLM 5.2 open-weight model for log …
Tracebit researchers found that planting prompt injections alongside secrets in Amazon Web Services reduced AI hacking agents' success rate from 57% to 5% for admin privilege escalation and from 36% t…
Kimi K3, the largest open-source model ever released by China's Moonshot AI, matches Claude's coding output in practice while costing a fraction of the price, according to a user's direct comparison. …
Databricks raised $3 billion at a $188 billion valuation, a 40% jump from its $134 billion mark five months ago, and revealed that its 3,000 engineers are defaulting to Chinese open-source model GLM 5…
Databricks announced a new funding round led by Coatue that values the company at $188 billion, extending its fundraising streak as it transforms into an AI provider. The round is expected to close la…
Chinese leader Xi Jinping called for more open-source AI and global regulation at the World Artificial Intelligence Conference in Shanghai on Friday, stating 'China is ready to be more open.' Xi outli…
Kilo users can now match skills to models by building custom agents that bundle a model, system prompt, and tool permissions into a single profile, enabling one-keystroke switching between agents like…