12:00
2026-10-05
tej.as
large-language-models
Kolibri vs Claude Sonnet 5.5: A German LLM Benchmark
Aleph Alpha's Kolibri lost all 136 blind comparisons against Claude Sonnet 5.5 in a German-language Bible deep-dive benchmark run by Dewfall developer on 3 October 2026, scoring 3.7 with German promptβ¦