cd/entity/BLEU· home› entities› BLEU
grep -l @bleu /news/*.json | wc -l → 5

BLEU

mentions 5 type Organization feed RSS

// recent coverage 5 mentions

16:24
2026-09-11
promptcube3.com
large-language-models

Why MCQ-based evaluation beats BLEU for video captions

A proposed evaluation method for video captioning replaces BLEU and METEOR with Multiple-Choice Question Answering (MCQA), scoring captions by the percentage of video-grounded questions an LLM can ans…

12:05
2026-08-05
pub.towardsai.net
artificial-intelligence

LLM-as-a-Judge: What It Is and How to Build One Yourself

LLM-as-a-Judge is a technique that uses a large language model to evaluate the output of another model, offering a scalable and explainable proxy for human preference. The approach was formalized in a…

04:40
2026-06-16
discuss.huggingface.co
large-language-models

Metrics for Text Generation from T5 Model

A user training a T5 model asked for alternative metrics to Exact Match for evaluating text generation. Community members suggested ROUGE-1, ROUGE-2, and BLEU, and recommended Braintrust for running e…

// co-occurs with top 8 entities
// topics top 6 topics