SOTAVerified

Transliteration

Transliteration is a mechanism for converting a word in a source (foreign) language to a target language, and often adopts approaches from machine translation. In machine translation, the objective is to preserve the semantic meaning of the utterance as much as possible while following the syntactic structure in the target language. In Transliteration, the objective is to preserve the original pronunciation of the source word as much as possible while following the phonological structures of the target language.

For example, the city’s name “Manchester” has become well known by people of languages other than English. These new words are often named entities that are important in cross-lingual information retrieval, information extraction, machine translation, and often present out-of-vocabulary challenges to spoken language technologies such as automatic speech recognition, spoken keyword search, and text-to-speech.

Source: Phonology-Augmented Statistical Framework for Machine Transliteration using Limited Linguistic Resources

Papers

Showing 151–175 of 435 papers

TitleStatusHype
Entity Clustering Across Languages—0
EPIK: Eliminating multi-model Pipelines with Knowledge-distillation—0
Bidirectional Bengali Script and Meetei Mayek Transliteration of Web Based Manipuri News Corpus—0
ANVITA Machine Translation System for WAT 2021 MultiIndicMT Shared Task—0
Exploiting Parallel Corpus for Handling Out-of-Vocabulary Words—0
Exploiting Transliterated Words for Finding Similarity in Inter-Language News Articles using Machine Learning—0
Exploring Linguistic Similarity and Zero-Shot Learning for Multilingual Translation of Dravidian Languages—0
Exploring the Role of Transliteration in In-Context Learning for Low-resource Languages Written in Non-Latin Scripts—0
Automatic Correction of Arabic Text: a Cascaded Approach—0
Factored Machine Translation Systems for Russian-English—0
False-Friend Detection and Entity Matching via Unsupervised Transliteration—0
Finite State Approach to the Kazakh Nominal Paradigm—0
Finite-state script normalization and processing utilities: The Nisaba Brahmic library—0
Foreign Words and the Automatic Processing of Arabic Social Media Text Written in Roman Script—0
Forward Transliteration of Dzongkha Text to Braille—0
Fourteen Light Tasks for comparing Analogical and Phrase-based Machine Translation—0
Digraph of Senegal s local languages: issues, challenges and prospects of their transliteration—0
Further Developments in Treebank Error Detection Using Derivation Trees—0
G2P Conversion of Proper Names Using Word Origin Information—0
Gender Prediction in English-Hindi Code-Mixed Social Media Content : Corpus and Baseline System—0
Graphonological Levenshtein Edit Distance: Application for Automated Cognate Identification—0
Digraphie des langues ouest africaines : Latin2Ajami : un algorithme de translitteration automatique—0
Gui at MixMT 2022 : English-Hinglish: An MT approach for translation of code mixed data—0
HCCL at SemEval-2017 Task 2: Combining Multilingual Word Embeddings and Transliteration Model for Semantic Similarity—0
A House United: Bridging the Script and Lexical Barrier between Hindi and Urdu—0
Show:102550
← PrevPage 7 of 18Next →

No leaderboard results yet.