SOTAVerified

Word Sense Disambiguation

The task of Word Sense Disambiguation (WSD) consists of associating words in context with their most suitable entry in a pre-defined sense inventory. The de-facto sense inventory for English in WSD is WordNet.. For example, given the word “mouse” and the following sentence:

“A mouse consists of an object held in one's hand, with one or more buttons.”

we would assign “mouse” with its electronic device sense (the 4th sense in the WordNet sense inventory).

Papers

Showing 1001–1035 of 1035 papers

TitleStatusHype
KPWr: Towards a Free Corpus of Polish—0
Evaluation of Classification Algorithms and Features for Collocation Extraction in Croatian—0
Mapping WordNet synsets to Wikipedia articles—0
Evaluating the Impact of Phrase Recognition on Concept Tagging—0
Assigning Connotation Values to Events—0
DBpedia: A Multilingual Cross-domain Knowledge Base—0
Is it Useful to Support Users with Lexical Resources? A User Study.—0
A new semantically annotated corpus with syntactic-semantic and cross-lingual senses—0
A Rough Set Formalization of Quantitative Evaluation with Ambiguity—0
Mapping WordNet to the Kyoto ontology—0
Bulgarian X-language Parallel Corpus—0
Using Verb Subcategorization for Word Sense Disambiguation—0
GerNED: A German Corpus for Named Entity Disambiguation—0
UBY-LMF -- A Uniform Model for Standardizing Heterogeneous Lexical-Semantic Resources in ISO-LMF—0
Using semi-experts to derive judgments on word sense alignment: a pilot study—0
Inforex -- a web-based tool for text corpus management and semantic annotation—0
A Comparative Evaluation of Word Sense Disambiguation Algorithms for German—0
Parallel Aligned Treebanks at LDC: New Challenges Interfacing Existing Infrastructures—0
DutchSemCor: Targeting the ideal sense-tagged corpus—0
Parallel Data, Tools and Interfaces in OPUS—0
Semi-Supervised Technical Term Tagging With Minimal User Feedback—0
Word Sense Inventories by Non-Experts.—0
Were the clocks striking or surprising? Using WSD to improve MT performance—0
Folheador: browsing through Portuguese semantic relations—0
WebCAGe -- A Web-Harvested Corpus Annotated with GermaNet Senses—0
UBY - A Large-Scale Unified Lexical-Semantic Resource Based on LMF—0
Subcat-LMF: Fleshing out a standardized format for subcategorization frame interoperability—0
Inferring Selectional Preferences from Part-Of-Speech N-grams—0
Multilingual Natural Language Processing—0
Looking at word meaning. An interactive visualization of Semantic Vector Spaces for Dutch synsets—0
Cross-Lingual Genre Classification—0
Word Sense Induction for Novel Sense Detection—0
Bootstrapping Events and Relations from Text—0
MaltOptimizer: An Optimization Tool for MaltParser—0
A Study of Hybrid Similarity Measures for Semantic Relation Extraction—0
Show:102550
← PrevPage 21 of 21Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1COSINE + Transductive LearningAccuracy85.3—Unverified
2PaLM 540B (finetuned)Accuracy78.8—Unverified
3ST-MoE-32B 269B (fine-tuned)Accuracy77.7—Unverified
4DeBERTa-EnsembleAccuracy77.5—Unverified
5Vega v2 6B (fine-tuned)Accuracy77.4—Unverified
6UL2 20B (fine-tuned)Accuracy77.3—Unverified
7Turing NLR v5 XXL 5.4B (fine-tuned)Accuracy77.1—Unverified
8T5-XXL 11BAccuracy76.9—Unverified
9DeBERTa-1.5BAccuracy76.4—Unverified
10ST-MoE-L 4.1B (fine-tuned)Accuracy74—Unverified
#ModelMetricClaimedVerifiedStatus
1SANDWiCHSenseval 287.8—Unverified
2GlossGPTSenseval 286.1—Unverified
3ConSeC+WNGCSenseval 282.7—Unverified
4ESR+WNGCSenseval 282.5—Unverified
5ConSeCSenseval 282.3—Unverified
6ESCHER SemCorSenseval 281.7—Unverified
7ESRSenseval 281.3—Unverified
8EWISER+WNGCSenseval 280.8—Unverified
9SemCor+WNGC, hypernymsSenseval 279.7—Unverified
10SparseLMMS+WNGCSenseval 279.6—Unverified
#ModelMetricClaimedVerifiedStatus
1Human BenchmarkAccuracy0.81—Unverified
2ruT5-large-finetuneAccuracy0.74—Unverified
3RuBERT conversationalAccuracy0.73—Unverified
4RuBERT plainAccuracy0.73—Unverified
5ruRoberta-large finetuneAccuracy0.72—Unverified
6ruBert-base finetuneAccuracy0.71—Unverified
7Multilingual BertAccuracy0.69—Unverified
8ruT5-base-finetuneAccuracy0.68—Unverified
9ruBert-large finetuneAccuracy0.68—Unverified
10SBERT_Large_mt_ru_finetuningAccuracy0.66—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF178.7—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF172.63—Unverified
3LSTMLP (T:SemCor, U:1K)F169.5—Unverified
4LSTMLP (T:OMSTI, U:1K)F168.1—Unverified
5LSTMLP (T:SemCor, U:OMSTI)F167.9—Unverified
6LSTM (T:OMSTI)F167.3—Unverified
7GASext (Concatenation)F167.2—Unverified
8GASext (Linear)F167.1—Unverified
9GAS (Concatenation)F167—Unverified
10LSTM (T:SemCor)F167—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF179.7—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF175.15—Unverified
3LSTMLP (T:OMSTI, U:1K)F174.4—Unverified
4LSTMLP (T:SemCor, U:OMSTI)F173.9—Unverified
5LSTMLP (T:SemCor, U:1K)F173.8—Unverified
6LSTM (T:SemCor)F173.6—Unverified
7GASext (Linear)F172.4—Unverified
8LSTM (T:OMSTI)F172.4—Unverified
9GASext (Concatenation)F172.2—Unverified
10GAS (Concatenation)F172.1—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF177.8—Unverified
2LSTMLP (T:SemCor, U:1K)F171.8—Unverified
3LSTMLP (T:SemCor, U:OMSTI)F171.1—Unverified
4LSTMLP (T:OMSTI, U:1K)F171—Unverified
5GASext (Concatenation)F170.5—Unverified
6GAS (Concatenation)F170.2—Unverified
7SemCor+WNGT, vocabulary reduced, ensembleF170.11—Unverified
8GASext (Linear)F170.1—Unverified
9GAS (Linear)F170—Unverified
10LSTM (T:SemCor)F169.2—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF190.4—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF186.02—Unverified
3kNN-BERT + POS (training corpus: WNGT)F185.32—Unverified
4LSTMLP (T:SemCor, U:OMSTI)F184.3—Unverified
5LSTMLP (T:SemCor, U:1K)F183.6—Unverified
6LSTMLP (T:OMSTI, U:1K)F183.3—Unverified
7LSTM (T:SemCor)F182.8—Unverified
8ShotgunWSD 2.0F181.22—Unverified
9kNN-BERTF181.2—Unverified
10LSTM (T:OMSTI)F181.1—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF173.4—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF166.81—Unverified
3LSTM (T:SemCor)F164.2—Unverified
4LSTMLP (T:SemCor, U:OMSTI)F163.7—Unverified
5LSTMLP (T:SemCor, U:1K)F163.5—Unverified
6LSTMLP (T:OMSTI, U:1K)F163.3—Unverified
7kNN-BERT + POS (training corpus: SemCor)F163.17—Unverified
8kNN-BERTF160.94—Unverified
9LSTM (T:OMSTI)F160.7—Unverified
#ModelMetricClaimedVerifiedStatus
1GlossGPTF1 (Zeroshot Dev)81.8—Unverified
2ESR LargeF1 (Zeroshot Dev)77.4—Unverified
3ESR baseF1 (Zeroshot Dev)73.9—Unverified
4SEMEq LargeF1 (Zeroshot Dev)73.7—Unverified
5SEMeq baseF1 (Zeroshot Dev)71.5—Unverified
6RTWE largeF1 (Zero shot test)69.9—Unverified
7LeskF1 (Zeroshot Dev)40.1—Unverified
8MFSF1 (Zeroshot Dev)0—Unverified
#ModelMetricClaimedVerifiedStatus
1HumanTask 3 Accuracy: all85.3—Unverified
2transformersTask 1 Accuracy: all77.8—Unverified
3CTLRTask 1 Accuracy: all76.8—Unverified
4GlossBert-wsTask 1 Accuracy: all75.9—Unverified
5Bert-baseTask 1 Accuracy: all75.3—Unverified
6Unsupervised BertTask 1 Accuracy: all54.4—Unverified
7FastTextTask 1 Accuracy: all53.7—Unverified
8All trueTask 1 Accuracy: all50.8—Unverified
#ModelMetricClaimedVerifiedStatus
1Chinchilla-70B (few-shot, k=5)Accuracy69.1—Unverified
2Gopher-280B (few-shot, k=5)Accuracy56.4—Unverified
3OPT 175BAccuracy49.1—Unverified
4GAL 120B (few-shot, k=5)Accuracy48.7—Unverified
5GAL 30B (few-shot, k=5)Accuracy47—Unverified
6BLOOM 176BAccuracy1.3—Unverified
#ModelMetricClaimedVerifiedStatus
1UKBppr_w2wSenseval 268.8—Unverified
2KEFAll68—Unverified
3WSD-TMAll66.9—Unverified
4BabelfyAll65.5—Unverified
5WN 1st sense baselineAll65.2—Unverified
6UKBppr_w2w-nfAll57.5—Unverified
#ModelMetricClaimedVerifiedStatus
1SemCor+WNGC, hypernymsF182.6—Unverified
2SemCor+WNGT, vocabulary reduced, ensembleF174.46—Unverified
3GASext (Concatenation)F172.6—Unverified
4GASext (Linear)F172.1—Unverified
5GAS (Concatenation)F171.8—Unverified
6GAS (Linear)F171.6—Unverified
#ModelMetricClaimedVerifiedStatus
1kNN-BERTF180.12—Unverified
2IMS + adapted CWF173.4—Unverified
3BiLSTM with GloVeF173.4—Unverified
4Single BiLSTMF172.5—Unverified
#ModelMetricClaimedVerifiedStatus
1kNN-BERTF176.52—Unverified
2BiLSTM with GloVeF166.9—Unverified
3IMS + adapted CWF166.2—Unverified
#ModelMetricClaimedVerifiedStatus
1SPINSequence Recovery %(All)30.3—Unverified