cd /news/large-language-models/typesafe-s-jev-independent-benchmark… · home › topics › large-language-models › article
[ARTICLE · art-140391] src=dev.to ↗ pub= topic=large-language-models verified=true sentiment=· neutral

TypeSafe's Jev: Independent Benchmark Against LLMs (with code)

A developer built an independent benchmark comparing TypeSafe's Jev model against GPT-4, Claude, and Gemini on classification tasks including spam detection, sentiment analysis, and topic classification, with reproducible code released on GitHub. Jev differs from conventional LLMs in that it outputs probabilities for given answer choices rather than generating text.

by read1 min views1 publishedSep 27, 2026

I built an independent benchmark to test TypeSafe's Jev model against GPT-4, Claude, and Gemini on classification tasks.

Jev is a different kind of model — instead of generating text, it outputs probabilities for given answer choices. This makes it particularly interesting for:

The benchmark covers spam detection, sentiment analysis, and topic classification tasks with full reproducible code.

Read the full article + code on Medium:

https://medium.com/@pravvich/typesafes-jev-beyond-the-hype-an-independent-benchmark-8bdc1c99d000GitHub repo:

https://github.com/PavelRavvich/jev-bench Connect with me on LinkedIn:

https://www.linkedin.com/in/pavel-ravvich/

── more in #large-language-models 4 stories · sorted by recency
── more on @typesafe 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/typesafe-s-jev-indep…] indexed:0 read:1min 2026-09-27 · —