SOTAVerified

Transliteration

Transliteration is a mechanism for converting a word in a source (foreign) language to a target language, and often adopts approaches from machine translation. In machine translation, the objective is to preserve the semantic meaning of the utterance as much as possible while following the syntactic structure in the target language. In Transliteration, the objective is to preserve the original pronunciation of the source word as much as possible while following the phonological structures of the target language.

For example, the city’s name “Manchester” has become well known by people of languages other than English. These new words are often named entities that are important in cross-lingual information retrieval, information extraction, machine translation, and often present out-of-vocabulary challenges to spoken language technologies such as automatic speech recognition, spoken keyword search, and text-to-speech.

Source: Phonology-Augmented Statistical Framework for Machine Transliteration using Limited Linguistic Resources

Papers

Showing 251–300 of 435 papers

TitleStatusHype
Regulating Orthography-Phonology Relationship for English to Thai Transliteration—0
Report of NEWS 2016 Machine Transliteration Shared Task—0
Substring-based unsupervised transliteration with phonetic and contextual knowledge—0
Egyptian Arabic to English Statistical Machine Translation System for NIST OpenMT'2015—0
A Correlational Encoder Decoder Architecture for Pivot Based Sequence Generation—0
Developing language technology tools and resources for a resource-poor language: Sindhi—0
Weighting Finite-State Transductions With Neural Context—0
Agreement on Target-bidirectional Neural Machine Translation—0
Morphological Analysis of Sahidic Coptic for Automatic Glossing—0
Arabic to English Person Name Transliteration using Twitter—0
Name Translation based on Fine-grained Named Entity Recognition in a Single Language—0
Data Cleaning for XML Electronic Dictionaries via Statistical Anomaly Detection—0
Mapping it differently: A solution to the linking challenges—0
Graphonological Levenshtein Edit Distance: Application for Automated Cognate Identification—0
Decoding Anagrammed Texts Written in an Unknown Language and Script—0
PJAIT Systems for the IWSLT 2015 Evaluation Campaign Enhanced by Comparable Corpora—0
POS Tagging of Hindi-English Code Mixed Text from Social Media: Some Machine Learning Experiments—0
Applying Sanskrit Concepts for Reordering in MT—0
An unsupervised EM method to infer time variation in sense probabilities—0
Translation of Unseen Bigrams by Analogy Using an SVM Classifier—0
amLite: Amharic Transliteration Using Key Map Dictionary—0
Training Automatic Transliteration Models on DBPedia Data—0
Improving Statistical Machine Translation with a Multilingual Paraphrase Database—0
Arabic Diacritization with Recurrent Neural Networks—0
Do we need bigram alignment models? On the effect of alignment quality on transduction accuracy in G2P—0
Improving Arabic Diacritization through Syntactic Analysis—0
Semi-supervised Chinese Word Segmentation based on Bilingual Information—0
Multiple System Combination for Transliteration—0
Neural Network Transduction Models in Transliteration Generation—0
Regularity and Flexibility in English-Chinese Name Transliteration—0
Boosting English-Chinese Machine Transliteration via High Quality Alignment and Multilingual Resources—0
Joint Arabic Segmentation and Part-Of-Speech Tagging—0
Classifying Arab Names Geographically—0
Whitepaper of NEWS 2015 Shared Task on Machine Transliteration—0
Data representation methods and use of mined corpora for Indian language transliteration—0
A Hybrid Transliteration Model for Chinese/English Named Entities ---BJTU-NLP Report for the 5th Named Entities Workshop—0
How do you spell that? A journey through word representations—0
Report of NEWS 2015 Machine Transliteration Shared Task—0
NCU IISR English-Korean and English-Chinese Named Entity Transliteration Using Different Grapheme Segmentation Approaches—0
AIDA2: A Hybrid Approach for Token and Sentence Level Dialect Identification in Arabic—0
Scalable Large-Margin Structured Learning: Theory and Algorithms—0
Structured Belief Propagation for NLP—0
An Empirical Study of Chinese Name Matching and ApplicationsCode0
Multiple Many-to-Many Sequence Alignment for Combining String-Valued Variables: A G2P Experiment—0
Lexicon Stratification for Translating Out-of-Vocabulary Words—0
Analyzing English-Spanish Named-Entity enhanced Machine Translation—0
What Matters Most in Morphologically Segmented SMT Models?—0
Brahmi-Net: A transliteration and script conversion system for languages of the Indian subcontinent—0
Automatic word stress annotation of Russian unrestricted text—0
Enhancing Sumerian Lemmatization by Unsupervised Named-Entity Recognition—0
Show:102550
← PrevPage 6 of 9Next →

No leaderboard results yet.