European Sovereign AI. Breakthrough Performance
SambaNova Systems launched a European sovereign AI inference platform that offers managed, dedicated, and on-premises deployment options with full data residency in German infrastructure. The platform…
SambaNova Systems launched a European sovereign AI inference platform that offers managed, dedicated, and on-premises deployment options with full data residency in German infrastructure. The platform…
Engineers can now run a full local agentic programming stack using Ollama, Gemma 4, and Claude Code, bypassing API costs, rate limits, and data privacy concerns. Google DeepMind's Gemma 4 26B MoE mode…
Google DeepMind released Quantization-Aware Training (QAT) checkpoints for the Gemma 4 model family, targeting local deployment on edge devices and consumer GPUs. The new Q4_0 QAT format reduces memor…
Google Cloud demonstrated a multi-cluster inference setup deploying an LLM across two regional GKE clusters using TPU v6e accelerators and GKE managed DRANET for networking. The experiment, documented…
Researchers have developed Soro, a family of Tajik-specialized conversational AI models built from Gemma 3 checkpoints and trained on a 1.9-billion-token Tajik corpus. The models outperform same-size …
Financial Times testing, conducted in partnership with AI safety group Alice, found that safety controls in Meta's Llama 3.3 and Google's Gemma 3 open-weight AI models can be removed in under 10 minut…
A 4-billion-parameter AI model released in early 2025 is now outperforming models seven times its size on standard reasoning benchmarks, with Google's Gemma 3 4B scoring 89.2% on GSM8K math reasoning …