SOTAVerified

Machine Translation

Machine translation is the task of translating a sentence in a source language to a different target language.

Approaches for machine translation can range from rule-based to statistical to neural-based. More recently, encoder-decoder attention-based architectures like BERT have attained major improvements in machine translation.

One of the most popular datasets used to benchmark machine translation systems is the WMT family of datasets. Some of the most commonly used evaluation metrics for machine translation systems include BLEU, METEOR, NIST, and others.

( Image credit: Google seq2seq )

Papers

Showing 351–400 of 10752 papers

TitleStatusHype
Dialectal and Low-Resource Machine Translation for Aromanian—0
Responsible Multilingual Large Language Models: A Survey of Development, Applications, and Societal Impact—0
Can General-Purpose Large Language Models Generalize to English-Thai Machine Translation ?—0
Subword Embedding from Bytes Gains Privacy without Sacrificing Accuracy and Complexity—0
Learning from others' mistakes: Finetuning machine translation models with span-level error annotations—0
On Creating an English-Thai Code-switched Machine Translation in Medical DomainCode0
Generalized Probabilistic Attention Mechanism in Transformers—0
Efficient Terminology Integration for LLM-based Translation in Specialized Domains—0
Analyzing Context Contributions in LLM-based Machine Translation—0
Grammatical Error Correction for Low-Resource Languages: The Case of Zarma—0
Back to School: Translation Using Grammar BooksCode0
mHumanEval -- A Multilingual Benchmark to Evaluate Large Language Models for Code GenerationCode0
Improving Vector-Quantized Image Modeling with Latent Consistency-Matching Diffusion—0
Analyzing Context Utilization of LLMs in Document-Level Translation—0
SwaQuAD-24: QA Benchmark Dataset in Swahili—0
Boosting LLM Translation Skills without General Ability Loss via Rationale Distillation—0
Towards Cross-Cultural Machine Translation with Retrieval-Augmented Generation from Multilingual Knowledge Graphs—0
NLIP_Lab-IITH Multilingual MT System for WAT24 MT Shared TaskCode0
Sarcasm Detection in a Less-Resourced LanguageCode0
"Is Hate Lost in Translation?": Evaluation of Multilingual LGBTQIA+ Hate Speech Detection—0
Enhancing Assamese NLP Capabilities: Introducing a Centralized Dataset RepositoryCode0
IntGrad MT: Eliciting LLMs' Machine Translation Capabilities with Sentence Interpolation and Gradual MT—0
PMMT: Preference Alignment in Multilingual Machine Translation via LLM Distillation—0
Watching the Watchers: Exposing Gender Disparities in Machine Translation Quality EstimationCode0
Effective Self-Mining of In-Context Examples for Unsupervised Machine Translation with LLMs—0
IsoChronoMeter: A simple and effective isochronic translation evaluation metricCode0
Beyond Human-Only: Evaluating Human-Machine Collaboration for Collecting High-Quality Translation Data—0
ChakmaNMT: A Low-resource Machine Translation On Chakma Language—0
Machine Translation Evaluation Benchmark for Wu Chinese: Workflow and Analysis—0
Code-Mixer Ya Nahi: Novel Approaches to Measuring Multilingual LLMs' Code-Mixing Capabilities—0
QE-EBM: Using Quality Estimators as Energy Loss for Machine Translation—0
Ukrainian-to-English folktale corpus: Parallel corpus creation and augmentation for machine translation in low-resource languages—0
State of NLP in Kenya: A Survey—0
SLAM-AAC: Enhancing Audio Captioning with Paraphrasing Augmentation and CLAP-Refine through LLMs—0
Adapters for Altering LLM Vocabularies: What Languages Benefit the Most?Code0
Balancing Innovation and Privacy: Data Security Strategies in Natural Language Processing Applications—0
Modeling User Preferences with Automatic Metrics: Creating a High-Quality Preference Dataset for Machine Translation—0
NusaMT-7B: Machine Translation for Low-Resource Indonesian Languages with Large Language Models—0
DelTA: An Online Document-Level Translation Agent Based on Multi-Level MemoryCode2
Mitigating the Language Mismatch and Repetition Issues in LLM-based Machine Translation via Model EditingCode0
Personal Intelligence System UniLM: Hybrid On-Device Small Language Model and Server-Based Large Language Model for Malay Nusantara—0
Are Large Language Models State-of-the-art Quality Estimators for Machine Translation of User-generated Content?Code0
Translation Canvas: An Explainable Interface to Pinpoint and Analyze Translation Systems—0
Neural machine translation system for Lezgian, Russian and Azerbaijani languagesCode0
On Instruction-Finetuning Neural Machine Translation Models—0
A test suite of prompt injection attacks for LLM-based machine translationCode0
Beyond Correlation: Interpretable Evaluation of Machine Translation Metrics—0
Leveraging Grammar Induction for Language Understanding and GenerationCode0
CTC-GMM: CTC guided modality matching for fast and accurate streaming speech translation—0
Toxic Subword Pruning for Dialogue Response Generation on Large Language Models—0
Show:102550
← PrevPage 8 of 216Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Transformer Cycle (Rev)BLEU score35.14—Unverified
2Noisy back-translationBLEU score35—Unverified
3Transformer+Rep(Uni)BLEU score33.89—Unverified
4T5-11BBLEU score32.1—Unverified
5BiBERTBLEU score31.26—Unverified
6Transformer + R-DropBLEU score30.91—Unverified
7Bi-SimCutBLEU score30.78—Unverified
8BERT-fused NMTBLEU score30.75—Unverified
9Data Diversification - TransformerBLEU score30.7—Unverified
10SimCutBLEU score30.56—Unverified
#ModelMetricClaimedVerifiedStatus
1Transformer+BT (ADMIN init)BLEU score46.4—Unverified
2Noisy back-translationBLEU score45.6—Unverified
3mRASP+Fine-TuneBLEU score44.3—Unverified
4Transformer + R-DropBLEU score43.95—Unverified
5AdminBLEU score43.8—Unverified
6Transformer (ADMIN init)BLEU score43.8—Unverified
7BERT-fused NMTBLEU score43.78—Unverified
8MUSE(Paralllel Multi-scale Attention)BLEU score43.5—Unverified
9T5BLEU score43.4—Unverified
10Local Joint Self-attentionBLEU score43.3—Unverified
#ModelMetricClaimedVerifiedStatus
1PiNMTBLEU score40.43—Unverified
2BiBERTBLEU score38.61—Unverified
3Bi-SimCutBLEU score38.37—Unverified
4Cutoff + Relaxed Attention + LMBLEU score37.96—Unverified
5DRDABLEU score37.95—Unverified
6Transformer + R-Drop + CutoffBLEU score37.9—Unverified
7SimCutBLEU score37.81—Unverified
8Cutoff+KneeBLEU score37.78—Unverified
9CutoffBLEU score37.6—Unverified
10CipherDAugBLEU score37.53—Unverified
#ModelMetricClaimedVerifiedStatus
1HWTSC-Teacher-SimScore19.97—Unverified
2MS-COMET-22Score19.89—Unverified
3MS-COMET-QE-22Score19.76—Unverified
4KG-BERTScoreScore17.28—Unverified
5metricx_xl_DA_2019Score17.17—Unverified
6COMET-QEScore16.8—Unverified
7COMET-22Score16.31—Unverified
8UniTE-srcScore15.68—Unverified
9UniTE-refScore15.38—Unverified
10metricx_xxl_DA_2019Score15.24—Unverified