SOTAVerified

Word Sense Disambiguation

The task of Word Sense Disambiguation (WSD) consists of associating words in context with their most suitable entry in a pre-defined sense inventory. The de-facto sense inventory for English in WSD is WordNet.. For example, given the word “mouse” and the following sentence:

“A mouse consists of an object held in one's hand, with one or more buttons.”

we would assign “mouse” with its electronic device sense (the 4th sense in the WordNet sense inventory).

Papers

Showing 276–300 of 1035 papers

TitleStatusHype
Estimating senses with sets of lexically related words for Polish word sense disambiguation—0
Compression de vocabulaire de sens gr\^ace aux relations s\'emantiques pour la d\'esambigu\" lexicale (Sense Vocabulary Compression through Semantic Knowledge for Word Sense Disambiguation)—0
Zero-shot Word Sense Disambiguation using Sense Definition EmbeddingsCode0
Just ``OneSeC'' for Producing Multilingual Sense-Annotated Data—0
Language Modelling Makes Sense: Propagating Representations through WordNet for Full-Coverage Word Sense DisambiguationCode1
LIAAD at SemDeep-5 Challenge: Word-in-Context (WiC)Code1
Making Fast Graph-based Algorithms with Graph Metric EmbeddingsCode0
L2F/INESC-ID at SemEval-2019 Task 2: Unsupervised Lexical Semantic Frame Induction using Contextualized Word Representations—0
Sense Vocabulary Compression through the Semantic Knowledge of WordNet for Neural Word Sense DisambiguationCode0
Using Wiktionary as a resource for WSD : the case of French verbs—0
In Search of Meaning: Lessons, Resources and Next Steps for Computational Analysis of Financial Discourse—0
Polylingual Wordnet—0
Fixed-Size Ordinally Forgetting Encoding Based Word Sense Disambiguation—0
Context based Analysis of Lexical Semantics for Hindi Language—0
A Contrastive Evaluation of Word Sense Disambiguation Systems for Finnish—0
Improving the Coverage and the Generalization Ability of Neural Word Sense Disambiguation through Hypernymy and Hyponymy Relationships—0
Local Homology of Word Embeddings—0
An Analysis of Attention Mechanisms: The Case of Word Sense Disambiguation in Neural Machine Translation—0
Integrating Weakly Supervised Word Sense Disambiguation into Neural Machine TranslationCode0
Sheffield Submissions for WMT18 Multimodal Translation Shared Task—0
The Word Sense Disambiguation Test Suite at WMT18—0
In-domain Context-aware Token Embeddings Improve Biomedical Named Entity Recognition—0
HUMIR at IEST-2018: Lexicon-Sensitive and Left-Right Context-Sensitive BiLSTM for Implicit Emotion Recognition—0
Similar but not the Same: Word Sense Disambiguation Improves Event Detection via Neural Representation Matching—0
Leveraging Gloss Knowledge in Neural Word Sense Disambiguation by Hierarchical Co-Attention—0
Show:102550
← PrevPage 12 of 42Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1COSINE + Transductive LearningAccuracy85.3—Unverified
2PaLM 540B (finetuned)Accuracy78.8—Unverified
3ST-MoE-32B 269B (fine-tuned)Accuracy77.7—Unverified
4DeBERTa-EnsembleAccuracy77.5—Unverified
5Vega v2 6B (fine-tuned)Accuracy77.4—Unverified
6UL2 20B (fine-tuned)Accuracy77.3—Unverified
7Turing NLR v5 XXL 5.4B (fine-tuned)Accuracy77.1—Unverified
8T5-XXL 11BAccuracy76.9—Unverified
9DeBERTa-1.5BAccuracy76.4—Unverified
10ST-MoE-L 4.1B (fine-tuned)Accuracy74—Unverified
#ModelMetricClaimedVerifiedStatus
1SANDWiCHSenseval 287.8—Unverified
2GlossGPTSenseval 286.1—Unverified
3ConSeC+WNGCSenseval 282.7—Unverified
4ESR+WNGCSenseval 282.5—Unverified
5ConSeCSenseval 282.3—Unverified
6ESCHER SemCorSenseval 281.7—Unverified
7ESRSenseval 281.3—Unverified
8EWISER+WNGCSenseval 280.8—Unverified
9SemCor+WNGC, hypernymsSenseval 279.7—Unverified
10SparseLMMS+WNGCSenseval 279.6—Unverified
#ModelMetricClaimedVerifiedStatus
1Human BenchmarkAccuracy0.81—Unverified
2ruT5-large-finetuneAccuracy0.74—Unverified
3RuBERT conversationalAccuracy0.73—Unverified
4RuBERT plainAccuracy0.73—Unverified
5ruRoberta-large finetuneAccuracy0.72—Unverified
6ruBert-base finetuneAccuracy0.71—Unverified
7Multilingual BertAccuracy0.69—Unverified
8ruT5-base-finetuneAccuracy0.68—Unverified
9ruBert-large finetuneAccuracy0.68—Unverified
10SBERT_Large_mt_ru_finetuningAccuracy0.66—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF178.7—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF172.63—Unverified
3LSTMLP (T:SemCor, U:1K)F169.5—Unverified
4LSTMLP (T:OMSTI, U:1K)F168.1—Unverified
5LSTMLP (T:SemCor, U:OMSTI)F167.9—Unverified
6LSTM (T:OMSTI)F167.3—Unverified
7GASext (Concatenation)F167.2—Unverified
8GASext (Linear)F167.1—Unverified
9GAS (Concatenation)F167—Unverified
10LSTM (T:SemCor)F167—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF179.7—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF175.15—Unverified
3LSTMLP (T:OMSTI, U:1K)F174.4—Unverified
4LSTMLP (T:SemCor, U:OMSTI)F173.9—Unverified
5LSTMLP (T:SemCor, U:1K)F173.8—Unverified
6LSTM (T:SemCor)F173.6—Unverified
7GASext (Linear)F172.4—Unverified
8LSTM (T:OMSTI)F172.4—Unverified
9GASext (Concatenation)F172.2—Unverified
10GAS (Concatenation)F172.1—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF177.8—Unverified
2LSTMLP (T:SemCor, U:1K)F171.8—Unverified
3LSTMLP (T:SemCor, U:OMSTI)F171.1—Unverified
4LSTMLP (T:OMSTI, U:1K)F171—Unverified
5GASext (Concatenation)F170.5—Unverified
6GAS (Concatenation)F170.2—Unverified
7SemCor+WNGT, vocabulary reduced, ensembleF170.11—Unverified
8GASext (Linear)F170.1—Unverified
9GAS (Linear)F170—Unverified
10LSTM (T:SemCor)F169.2—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF190.4—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF186.02—Unverified
3kNN-BERT + POS (training corpus: WNGT)F185.32—Unverified
4LSTMLP (T:SemCor, U:OMSTI)F184.3—Unverified
5LSTMLP (T:SemCor, U:1K)F183.6—Unverified
6LSTMLP (T:OMSTI, U:1K)F183.3—Unverified
7LSTM (T:SemCor)F182.8—Unverified
8ShotgunWSD 2.0F181.22—Unverified
9kNN-BERTF181.2—Unverified
10LSTM (T:OMSTI)F181.1—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF173.4—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF166.81—Unverified
3LSTM (T:SemCor)F164.2—Unverified
4LSTMLP (T:SemCor, U:OMSTI)F163.7—Unverified
5LSTMLP (T:SemCor, U:1K)F163.5—Unverified
6LSTMLP (T:OMSTI, U:1K)F163.3—Unverified
7kNN-BERT + POS (training corpus: SemCor)F163.17—Unverified
8kNN-BERTF160.94—Unverified
9LSTM (T:OMSTI)F160.7—Unverified
#ModelMetricClaimedVerifiedStatus
1GlossGPTF1 (Zeroshot Dev)81.8—Unverified
2ESR LargeF1 (Zeroshot Dev)77.4—Unverified
3ESR baseF1 (Zeroshot Dev)73.9—Unverified
4SEMEq LargeF1 (Zeroshot Dev)73.7—Unverified
5SEMeq baseF1 (Zeroshot Dev)71.5—Unverified
6RTWE largeF1 (Zero shot test)69.9—Unverified
7LeskF1 (Zeroshot Dev)40.1—Unverified
8MFSF1 (Zeroshot Dev)0—Unverified
#ModelMetricClaimedVerifiedStatus
1HumanTask 3 Accuracy: all85.3—Unverified
2transformersTask 1 Accuracy: all77.8—Unverified
3CTLRTask 1 Accuracy: all76.8—Unverified
4GlossBert-wsTask 1 Accuracy: all75.9—Unverified
5Bert-baseTask 1 Accuracy: all75.3—Unverified
6Unsupervised BertTask 1 Accuracy: all54.4—Unverified
7FastTextTask 1 Accuracy: all53.7—Unverified
8All trueTask 1 Accuracy: all50.8—Unverified
#ModelMetricClaimedVerifiedStatus
1Chinchilla-70B (few-shot, k=5)Accuracy69.1—Unverified
2Gopher-280B (few-shot, k=5)Accuracy56.4—Unverified
3OPT 175BAccuracy49.1—Unverified
4GAL 120B (few-shot, k=5)Accuracy48.7—Unverified
5GAL 30B (few-shot, k=5)Accuracy47—Unverified
6BLOOM 176BAccuracy1.3—Unverified
#ModelMetricClaimedVerifiedStatus
1UKBppr_w2wSenseval 268.8—Unverified
2KEFAll68—Unverified
3WSD-TMAll66.9—Unverified
4BabelfyAll65.5—Unverified
5WN 1st sense baselineAll65.2—Unverified
6UKBppr_w2w-nfAll57.5—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF182.6—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF174.46—Unverified
3GASext (Concatenation)F172.6—Unverified
4GASext (Linear)F172.1—Unverified
5GAS (Concatenation)F171.8—Unverified
6GAS (Linear)F171.6—Unverified
#ModelMetricClaimedVerifiedStatus
1kNN-BERTF180.12—Unverified
2IMS + adapted CWF173.4—Unverified
3BiLSTM with GloVeF173.4—Unverified
4Single BiLSTMF172.5—Unverified
#ModelMetricClaimedVerifiedStatus
1kNN-BERTF176.52—Unverified
2BiLSTM with GloVeF166.9—Unverified
3IMS + adapted CWF166.2—Unverified
#ModelMetricClaimedVerifiedStatus
1SPINSequence Recovery %(All)30.3—Unverified