04:00
2026-10-08
arxiv.org
large-language-models
sk-bench: A Native-First Benchmark for Evaluating Large Language Models in Slovak
A new arXiv paper (2610.09152v1) introduces sk-bench, a native-first Slovak benchmark comprising 30 datasets and 33 scored task variants across ten skill categories, and reports that the best open-wei…