SOTAVerified

Entity Resolution

Entity resolution (also known as entity matching, record linkage, or duplicate detection) is the task of finding records that refer to the same real-world entity across different data sources (e.g., data files, books, websites, and databases). (Source: Wikipedia)

Surveys on entity resolution:

The task of entity resolution is closely related to the task of entity alignment which focuses on matching entities between knowledge bases. The task of entity linking differs from entity resolution as entity linking focuses on identifying entity mentions in free text.

Papers

Showing 101–125 of 184 papers

TitleStatusHype
Clustering with Fast, Automated and Reproducible assessment applied to longitudinal neural tracking—0
The STEM-ECR Dataset: Grounding Scientific Entity References in STEM Scholarly Content to Authoritative Encyclopedic and Lexicographic Sources—0
Crowdsourced Collective Entity Resolution with Relational Match PropagationCode0
Pre-Training for Query Rewriting in A Spoken Language Understanding System—0
AutoBlock: A Hands-off Blocking Framework for Entity MatchingCode1
Towards Interpretable and Learnable Risk Analysis for Entity Resolution—0
Entity resolution for noisy ASR transcripts—0
d-blink: Distributed End-to-End Bayesian Entity ResolutionCode0
Accelerating Column Generation via Flexible Dual Optimal Inequalities with Application to Entity ResolutionCode0
Learning to Sample: an Active Learning Framework—0
Local Embeddings for Relational Data Integration—0
ZeroER: Entity Resolution using Zero Labeled ExamplesCode0
The Swedish PoliGraph: A Semantic Graph for Argument Mining of Swedish Parliamentary Data—0
Optimal Transport-based Alignment of Learned Character Representations for String SimilarityCode0
Crowdsourcing and Aggregating Nested Markable AnnotationsCode0
Low-resource Deep Entity Resolution with Transfer and Active Learning—0
Knowledge Refinement via Rule Selection—0
Topological and Semantic Graph-based Author Disambiguation on DBLP Data in Neo4j—0
NSEEN: Neural Semantic Embedding for Entity Normalization—0
Integrating User Feedback under Identity Uncertainty in Knowledge Base Construction—0
Gradual Machine Learning for Entity Resolution—0
Probabilistic Blocking with An Application to the Syrian Conflict—0
A Practical Approach to Proper Inference with Linked Data—0
Scalable Matching and Clustering of Entities with FAMER—0
Learning Text Representations for 500K Classification Tasks on Named Entity DisambiguationCode0
Show:102550
← PrevPage 5 of 8Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1gpt4-0613_fewshot-10F1 (%)85.21—Unverified
2gpt-4o-mini-2024-07-18_fine_tunedF1 (%)80.25—Unverified
3RoBERTa-SupConF1 (%)79.28—Unverified
4RobEMF1 (%)79.06—Unverified
5Random ForestF1 (%)79—Unverified
6HGF1 (%)76.4—Unverified
7DittoF1 (%)75.58—Unverified
8CorDEL-SumF1 (%)70.2—Unverified
9DeepMatcher - HybridF1 (%)69.3—Unverified
10D-HATF1 (%)67.5—Unverified
#ModelMetricClaimedVerifiedStatus
1gpt4-0613_zeroshotF1 (%)95.78—Unverified
2RoBERTa-SupConF1 (%)94.29—Unverified
3gpt-4o-mini-2024-07-18_fine_tunedF1 (%)94.09—Unverified
4gpt-4o-2024-08-06F1 (%)92.2—Unverified
5RobEMF1 (%)90.9—Unverified
6HGF1 (%)89.8—Unverified
7DittoF1 (%)89.33—Unverified
8gpt-4o-mini-2024-07-18F1 (%)87.68—Unverified
9Meta-Llama-3.1-8B-Instruct_fine_tunedF1 (%)87.34—Unverified
10Random ForestF1 (%)85—Unverified
#ModelMetricClaimedVerifiedStatus
1gpt4-0613_zeroshotF1 (%)89.61—Unverified
2gpt-4o-2024-08-06_fine_tuned_wdc_smallF1 (%)87.1—Unverified
3gpt-4o-mini-2024-07-18_structured_explanationsF1 (%)84.38—Unverified
4gpt-4o-mini-2024-07-18F1 (%)81.61—Unverified
5RoBERTa-SupConF1 (%)79.99—Unverified
6Llama3.1_70B_structured_explanationsF1 (%)76.7—Unverified
7Llama3.1_70BF1 (%)75.2—Unverified
8Llama3.1_8B_error-based_example_selectionF1 (%)74.37—Unverified
9Llama3.1_8B_structured_explanationsF1 (%)74.13—Unverified
10DittoF1 (%)73.93—Unverified
#ModelMetricClaimedVerifiedStatus
1BERTF1 (%)96.53—Unverified
2RoBERTa-SupConF1 (%)95.21—Unverified
3HGF1 (%)88.5—Unverified
4DADER-MMDF1 (%)88—Unverified
5DittoF1 (%)80.76—Unverified
6JointBERTF1 (%)77.55—Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa-SupConF1 (%)98.33—Unverified
2JointBERTF1 (%)97.49—Unverified
3BERTF1 (%)97.37—Unverified
4HGF1 (%)96.5—Unverified
5DittoF1 (%)95.45—Unverified
6Random ForestF1 (%)78—Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa-baseF1 (%)71.14—Unverified
2DittoF1 (%)70.66—Unverified
3HGF1 (%)68.74—Unverified
4RoBERTa-SupConF1 (%)57.23—Unverified
#ModelMetricClaimedVerifiedStatus
1HGF1 (%)94—Unverified
2DADER-NoDAF1 (%)88.6—Unverified
3DittoF1 (%)85.12—Unverified
4JointBERTF1 (%)75.83—Unverified
#ModelMetricClaimedVerifiedStatus
1ALMSER-GBF10.95—Unverified
2FAMER-SplitMergeF10.88—Unverified
3FAMER-SplitF10.84—Unverified
#ModelMetricClaimedVerifiedStatus
1JointBERTF1 (%)97.09—Unverified
2DittoF1 (%)96.53—Unverified
3HGF1 (%)96.5—Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa-SupConF1 Micro88.63—Unverified
2RoBERTa-baseF1 Micro52.03—Unverified
#ModelMetricClaimedVerifiedStatus
1gpt-4o-2024-08-06_fine_tuned_wdc_smallF1 (%)87.07—Unverified