Apertus Mini
Swiss AI researchers released the Apertus Mini collection, 16 small language models distilled from the Apertus v1 8B model, available in 0.5B, 1.5B, and 4B parameter sizes with multiple quantization lβ¦
Swiss AI researchers released the Apertus Mini collection, 16 small language models distilled from the Apertus v1 8B model, available in 0.5B, 1.5B, and 4B parameter sizes with multiple quantization lβ¦
On June 11, 2026, OpenAI committed to a 10-gigawatt data center, marking a shift toward national-scale AI infrastructure, while a separate study showed that finetuning a model on a book's plot alone rβ¦
A small study testing whether wrapping untrusted prompt inputs in mock tool calls could improve language model robustness found the technique did not broadly help across three LLM-as-a-Judge tasks andβ¦
Researchers from NVIDIA have developed a new sparse data format and custom GPU kernels, called TwELL, that reshape unstructured sparsity in transformer language models to align with GPU architecture, β¦