SOTAVerified

Chunking

Chunking, also known as shallow parsing, identifies continuous spans of tokens that form syntactic units such as noun phrases or verb phrases.

Example:

| Vinken | , | 61 | years | old | | --- | ---| --- | --- | --- | | B-NLP| I-NP | I-NP | I-NP | I-NP |

Papers

Showing 26–50 of 447 papers

TitleStatusHype
S2 Chunking: A Hybrid Framework for Document Segmentation Through Integrated Spatial and Semantic AnalysisCode1
Sparse Modular Activation for Efficient Sequence ModelingCode1
Problem Solved? Information Extraction Design Space for Layout-Rich Documents using LLMsCode1
TimeLoc: A Unified End-to-End Framework for Precise Timestamp Localization in Long VideosCode1
Recurrent Chunking Mechanisms for Long-Text Machine Reading ComprehensionCode1
NeuSym-RAG: Hybrid Neural Symbolic Retrieval with Multiview Structuring for PDF Question AnsweringCode1
On LLM-Enhanced Mixed-Type Data Imputation with High-Order Message PassingCode1
Learning Variable Compliance Control From a Few Demonstrations for Bimanual Robot with Haptic Feedback Teleoperation SystemCode1
Capturing Global Informativeness in Open Domain Keyphrase ExtractionCode1
Improving Named Entity Recognition by External Context Retrieving and Cooperative LearningCode1
Optimal Hyperparameters for Deep LSTM-Networks for Sequence Labeling TasksCode1
Dataset Decomposition: Faster LLM Training with Variable Sequence Length CurriculumCode1
Fast and Accurate Factual Inconsistency Detection Over Long DocumentsCode1
Context is Gold to find the Gold Passage: Evaluating and Training Contextual Document EmbeddingsCode1
CoFE-RAG: A Comprehensive Full-chain Evaluation Framework for Retrieval-Augmented Generation with Enhanced Data DiversityCode1
AIN: Fast and Accurate Sequence Labeling with Approximate Inference NetworkCode1
Leveraging Fine-Tuned Retrieval-Augmented Generation with Long-Context Support: For 3GPP StandardsCode1
Chat3GPP: An Open-Source Retrieval-Augmented Generation Framework for 3GPP DocumentsCode1
Automated Concatenation of Embeddings for Structured PredictionCode1
BERTraffic: BERT-based Joint Speaker Role and Speaker Change Detection for Air Traffic Control CommunicationsCode1
NetKet 3: Machine Learning Toolbox for Many-Body Quantum SystemsCode1
ChordMixer: A Scalable Neural Attention Model for Sequences with Different LengthsCode1
Paradigm Shift in Natural Language ProcessingCode1
An Experimental Investigation of Part-Of-Speech Taggers for Vietnamese—0
An Experimental Comparison of Active Learning Strategies for Partially Labeled Sequences—0
Show:102550
← PrevPage 2 of 18Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1ACEExact Span F197.3—Unverified
2BERT-CRF (Replicated in AdaSeq)Exact Span F197.18—Unverified
3ELMo + MAT + Multi-TaskExact Span F197.04—Unverified
4CVT+Multi-Task+LargeExact Span F196.98—Unverified
5ELMo + Multi-TaskExact Span F196.83—Unverified
6FlairExact Span F196.72—Unverified
7SeqVATExact Span F195.45—Unverified
8Adversarial TrainingExact Span F195.25—Unverified
9BiLSTM-CRFExact Span F195.18—Unverified
#ModelMetricClaimedVerifiedStatus
1ACEF1 score97.3—Unverified
2Flair embeddingsF1 score96.72—Unverified
3JMTF1 score95.77—Unverified
4Low supervisionF1 score95.57—Unverified
5IntNet + BiLSTM-CRFF1 score95.29—Unverified
6Suzuki and IsozakiF1 score95.15—Unverified
7NCRF++F1 score95.06—Unverified
8BI-LSTM-CRF (Senna) (ours)F1 score94.46—Unverified
#ModelMetricClaimedVerifiedStatus
1ACEF195—Unverified
2Wang et al., 2020F194.4—Unverified
3AINF194.04—Unverified
#ModelMetricClaimedVerifiedStatus
1Wang et al., 2020F192—Unverified
2AINF191.71—Unverified
#ModelMetricClaimedVerifiedStatus
1Def2VecAUC93.07—Unverified