15:00
2026-07-30
akitaonrails.com
artificial-intelligence
Novo LLM Benchmark: refiz todos os testes!
Fabio Akita released version 2 of his LLM Coding Benchmark, reporting that scores are not directly comparable to version 1 due to changes in prompts, requirements, harnesses, validation, and rubric. Tโฆ