SOTAVerified

Machine Translation

Machine translation is the task of translating a sentence in a source language to a different target language.

Approaches for machine translation can range from rule-based to statistical to neural-based. More recently, encoder-decoder attention-based architectures like BERT have attained major improvements in machine translation.

One of the most popular datasets used to benchmark machine translation systems is the WMT family of datasets. Some of the most commonly used evaluation metrics for machine translation systems include BLEU, METEOR, NIST, and others.

( Image credit: Google seq2seq )

Papers

Showing 1–10 of 10752 papers

TitleStatusHype
Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings—0
Speak2Sign3D: A Multi-modal Pipeline for English Speech to American Sign Language Animation—0
GRAFT: A Graph-based Flow-aware Agentic Framework for Document-level Machine Translation—0
TransLaw: Benchmarking Large Language Models in Multi-Agent Simulation of the Collaborative Translation—0
Enhancing Automatic Term Extraction with Large Language Models via Syntactic Retrieval—0
Intrinsic vs. Extrinsic Evaluation of Czech Sentence Embeddings: Semantic Relevance Doesn't Help with MT Evaluation—0
Has Machine Translation Evaluation Achieved Human Parity? The Human Reference and the Limits of ProgressCode0
CycleDistill: Bootstrapping Machine Translation using LLMs with Cyclical DistillationCode0
Semantic similarity estimation for domain specific data using BERT and other techniques—0
Sequence-to-Sequence Models with Attention Mechanistically Map to the Architecture of Human Memory Search—0
Show:102550
← PrevPage 1 of 1076Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Bi-SimCutBLEU score35.15—Unverified
2BiBERTBLEU score34.94—Unverified
3SimCutBLEU score34.86—Unverified
4MegaBLEU score33.12—Unverified
5CMLM+LAT+4 iterationsBLEU score32.04—Unverified
6MAT+KneeBLEU score31.9—Unverified
7CNATBLEU score30.75—Unverified
8CMLM+LAT+1 iterationsBLEU score29.91—Unverified
9FlowSeq-large (NPD n = 30)BLEU score28.29—Unverified
10FlowSeq-large (NPD n = 15)BLEU score27.71—Unverified