BenchVibe AI Ecosystem

VIP 👤

🏠 Beranda

Benchmark

📊 Semua Benchmark 🦖 Dinosaurus v1 🦖 Dinosaurus v2 ✅ Aplikasi To-Do List 🎨 Halaman Bebas Kreatif 🎯 FSACB - Showcase Utama 🌍 Benchmark Terjemahan

Model

🏆 Top 10 Model 🆓 Model Gratis 📋 Semua Model ⚙️ Kilo Code

Sumber Daya

💬 Perpustakaan Prompt 📖 Glosarium AI 🔗 Tautan Berguna

📖

Evaluation and Metrics

BLEU (Bilingual Evaluation Understudy)

Automatic metric for evaluating the quality of machine translations by comparing the n-gram precision of the generated text against one or more human references. It measures the overlap of text segments between the model output and the reference.

← Kembali