Home Glossary BLEU Score

BLEU Score - Page 70

BLEU, or Bilingual Evaluation Understudy, is an automatic metric originally developed for machine translation. It compares short sequences of words in generated text with one or more reference translations, combines precision across several n-gram lengths, and applies a penalty when output is too short. BLEU is inexpensive and useful for comparing systems on the same dataset, but it does not directly measure factuality, fluency, meaning, or human preference. Valid alternative wording can receive a low score, while awkward text can match many reference phrases. Results depend on tokenization and implementation, so reporting should include the exact evaluation settings and, when possible, human judgment.

Andrew Tulloch Leaves $12B AI Startup to Join Meta After Turning Down $1.5B Offer
By • 3 mins read
AI & Machine Learning, Immersive Reality (AR, VR, MR, and XR), News, Startups & Investment

Andrew Tulloch Leaves $12B AI Startup to Join Meta After Turning Down $1.5B Offer

By • 3 mins read

Andrew Tulloch, co-founder of the $12 billion AI startup Thinking Machines Lab, has joined Meta after previously rejecting what reports described as a $1.5 billion offer — a figure Meta has since called ‘inaccurate and ridiculous.’