BenchMark·Hub

NTEU_Multilingual_Evaluation_Dataset

B待确认Benchmark
语言/语义翻译开放cc-by-4.0

发布方:BSC-LT

热度10.7±0
下载量 · 30天
157
Hugging Face
GitHub Stars
代码仓库
论文被引
Semantic Scholar
跑分模型 · 30天
Leaderboard results

简介

Dataset Card for NTEU Multilingual Evaluation Dataset Dataset Summary This evaluation dataset for Machine Translation was created by the NTEU - Neural Translation for the EU project. The evaluation dataset includes around 1,000 parallel sentences in the 24 official European languages. The original NTEU dataset has been cleaned and filtered by removing empty lines and near-duplicates, and it has been augmented with Catalan. The Catalan version was manually produced by a… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/NTEU_Multilingual_Evaluation_Dataset.