Introducing Gemma 4 models on Amazon Bedrock
Amazon Bedrock announced the availability of Gemma 4 models, a family of open-weight AI models from Google DeepMind, including dense and mixture-of-experts variants with built-in reasoning, function c…
Amazon Bedrock announced the availability of Gemma 4 models, a family of open-weight AI models from Google DeepMind, including dense and mixture-of-experts variants with built-in reasoning, function c…
Artificial Analysis released AgentPerf, the first benchmark for agentic-AI infrastructure, which replays recorded multi-step agent trajectories instead of single chat completions. NVIDIA reported that…
NVIDIA released Nemotron 3 Ultra, a 550-billion-parameter open-weights reasoning model, on June 4, 2026. It is the best US open model by Artificial Analysis's scoring but trails Chinese leader Kimi K2…
Artificial Analysis released its Frontier Language Model Intelligence index, tracking performance, cost, and execution time of leading AI models over time. The index evaluates models on agentic tasks,…
NVIDIA's Blackwell platform achieved up to 20x more agents per megawatt than the previous generation in the first AgentPerf benchmark, a new test from independent firm Artificial Analysis designed for…
Artificial Analysis, an independent AI benchmarking platform, launched its Coding Agent Benchmarks and Index at a June 11 event in San Francisco featuring speakers from Cognition, Cursor, and NVIDIA. …
Artificial Analysis released initial results for its AA-AgentPerf benchmark, showing NVIDIA's Blackwell systems outperforming AMD's Instinct MI355X GPUs on power-efficient agentic inference using Deep…
NVIDIA's Blackwell architecture leads the inaugural AgentPerf benchmark for agentic AI. AMD opens pre-orders for its Linux-friendly Ryzen AI Halo developer platform featuring the Strix Halo APU. Linux…
Nvidia's Blackwell architecture achieves 20 times more AI agents per megawatt than its Hopper generation, according to the new AgentPerf benchmark. The efficiency leap, driven by FP4 precision and adv…
NVIDIA achieved leading agentic coding performance on the first agentic AI benchmark, AA-AgentPerf, delivering up to 20x better performance than previous generations. The benchmark, created by Artific…
NVIDIA's Blackwell Ultra NVL72 platform achieved top performance on the first agentic AI benchmark, AgentPerf from Artificial Analysis, running up to 20 times more agents per megawatt than the previou…
Fable 5 has achieved a performance level on par with GPT-5.5 in the Artificial Analysis Coding Agent Index, a composite benchmark measuring real-world coding agent performance across software engineer…
NVIDIA released Nemotron 3 Ultra, a 550-billion-parameter open-weight model, on June 1, 2026, and it became available on Ollama's cloud three days later. The model uses a sparse Mixture-of-Experts des…
Cohere launched North Mini Code, an open-source mixture-of-experts model with 30B total parameters and 3B active, optimized for agentic coding tasks. Released under Apache 2.0, it aims to provide deve…
NVIDIA released Nemotron 3.5 ASR, a 600M-parameter streaming multilingual speech-to-text model that transcribes 40 language-locales from a single checkpoint with built-in punctuation and capitalizatio…
NVIDIA released Nemotron 3 Ultra, a 550-billion-parameter open coding model with 55 billion active parameters per token, designed for fast, agentic developer workflows. The model achieves over 300 out…
Nvidia announced that its Cosmos 3 model for physical AI ranks first on seven leaderboards for world generation, robot action policy, and industrial vision understanding. The company described Cosmos …
NVIDIA researchers introduced Cosmos 3, a family of omnimodal world models that jointly process and generate language, image, video, audio, and action sequences within a unified mixture-of-transformer…
Microsoft added average token usage to a model release card yesterday, introducing a new dual benchmark that measures both performance and cost. The metric shows Microsoft's model achieving a 71.6 SWE…
Claude Opus 4.8 is now available on AWS Bedrock for enterprise AI workloads and as a generally available model for GitHub Copilot, placing the model directly in production infrastructure and developer…