SOTAVerified

Transliteration

Transliteration is a mechanism for converting a word in a source (foreign) language to a target language, and often adopts approaches from machine translation. In machine translation, the objective is to preserve the semantic meaning of the utterance as much as possible while following the syntactic structure in the target language. In Transliteration, the objective is to preserve the original pronunciation of the source word as much as possible while following the phonological structures of the target language.

For example, the city’s name “Manchester” has become well known by people of languages other than English. These new words are often named entities that are important in cross-lingual information retrieval, information extraction, machine translation, and often present out-of-vocabulary challenges to spoken language technologies such as automatic speech recognition, spoken keyword search, and text-to-speech.

Source: Phonology-Augmented Statistical Framework for Machine Transliteration using Limited Linguistic Resources

Papers

Showing 301–350 of 435 papers

TitleStatusHype
Automatic word stress annotation of Russian unrestricted text—0
Joint Generation of Transliterations from Multiple Representations—0
Enhancing Sumerian Lemmatization by Unsupervised Named-Entity Recognition—0
Model Invertibility Regularization: Sequence Alignment With or Without Parallel Data—0
``ye word kis lang ka hai bhai?'' Testing the Limits of Word level Language Identification—0
English to Punjabi Transliteration using Orthographic and Phonetic Information—0
HinMA: Distributed Morphology based Hindi Morphological Analyzer—0
Sangam: A Perso-Arabic to Indic Script Machine Transliteration Model—0
Word-level Language Identification in Bi-lingual Code-switched Texts—0
Automatic Correction of Arabic Text: a Cascaded Approach—0
Foreign Words and the Automatic Processing of Arabic Social Media Text Written in Roman Script—0
A Framework for the Classification and Annotation of Multiword Expressions in Dialectal Arabic—0
POS Tagging of English-Hindi Code-Mixed Social Media Content—0
``I am borrowing ya mixing ?'' An Analysis of English-Hindi Code Mixing in Facebook—0
Transliteration of Arabizi into Arabic Orthography: Developing a Parallel Annotated Arabizi-Arabic Script SMS/Chat Corpus—0
Proper Name Machine Translation from Japanese to Japanese Sign Language—0
Transliteration Extraction from Classical Chinese Buddhist Literature Using Conditional Random Fields with Language Models—0
Solving Substitution Ciphers with Combined Language Models—0
Konkanverter - A Finite State Transducer based Statistical Machine Transliteration Engine for Konkani Language—0
3arif: A Corpus of Modern Standard and Egyptian Arabic Tweets Annotated for Epistemic Modality Using Interactive Crowdsourcing—0
Confusion Network for Arabic Name Disambiguation and Transliteration in Statistical Machine Translation—0
Fourteen Light Tasks for comparing Analogical and Phrase-based Machine Translation—0
Assamese-English Bilingual Machine Translation—0
Study of the impact of proper name transliteration on the performance of word alignment in French-Arabic parallel corpora (Etude de l'impact de la translitt\'eration de noms propres sur la qualit\'e de l'alignement de mots \`a partir de corpus parall\`eles fran -arabe) [in French]—0
How to Speak a Language without Knowing It—0
DCU Terminology Translation System for Medical Query Subtask at WMT14—0
Edinburgh's Syntax-Based Systems at WMT 2014—0
Yandex School of Data Analysis Russian-English Machine Translation System for WMT14—0
Automatic Transliteration of Romanized Dialectal Arabic—0
Stochastic Contextual Edit Distance and Probabilistic FSTs—0
AraNLP: a Java-based Library for the Processing of Arabic Text.—0
Automatic acquisition of Urdu nouns (along with gender and irregular plurals)—0
Shata-Anuvadak: Tackling Multiway Translation of Indian Languages—0
Vocabulary-Based Language Similarity using Web Corpora—0
Automatic detection of other-repetition occurrences: application to French conversational Speech—0
When Transliteration Met Crowdsourcing : An Empirical Study of Transliteration via Crowdsourcing using Efficient, Non-redundant and Fair Quality Control—0
Bilingual dictionaries for all EU languagesCode0
Transliteration and alignment of parallel texts from Cyrillic to Latin—0
Bilingual Dictionary Construction with Transliteration Filtering—0
A Conventional Orthography for Tunisian Arabic—0
MADAMIRA: A Fast, Comprehensive Tool for Morphological Analysis and Disambiguation of Arabic—0
Map Translation Using Geo-tagged Social Media—0
Integrating an Unsupervised Transliteration Model into Statistical Machine Translation—0
Hybrid Approach to English-Hindi Name Entity Transliteration—0
Improving Statistical Machine Translation for a Resource-Poor Language Using Related Resource-Rich Languages—0
Improving Performance Of English-Hindi Cross Language Information Retrieval Using Transliteration Of Query Terms—0
Assamese WordNet based Quality Enhancement of Bilingual Machine Translation System—0
Transliteration Extraction from Classical Chinese Buddhist Literature Using Conditional Random Fields—0
Exploiting Parallel Corpus for Handling Out-of-Vocabulary Words—0
Transliteration Systems across Indian Languages Using Parallel Corpora—0
Show:102550
← PrevPage 7 of 9Next →

No leaderboard results yet.