SOTAVerified is in maintenance mode — catalog snapshot as of 2026-10-10. Details
SOTAVerified

Natural Language Understanding

Natural Language Understanding is an important field of Natural Language Processing which contains various tasks such as text classification, natural language inference and story comprehension. Applications enabled by natural language understanding range from question answering to automated reasoning.

Source: Find a Reasonable Ending for Stories: Does Logic Relation Help the Story Cloze Test?

Papers

Showing 1–12 of 12 papers

TitleStatusHype
Mixture-of-Agents Enhances Large Language Model CapabilitiesCode7
DataComp-LM: In search of the next generation of training sets for language modelsCode7
Harnessing the Power of LLMs in Practice: A Survey on ChatGPT and BeyondCode6
Agentic Retrieval-Augmented Generation: A Survey on Agentic RAGCode5
MedCare: Advancing Medical LLMs through Decoupling Clinical Alignment and Knowledge AggregationCode5
MING-MOE: Enhancing Medical Multi-Task Learning in Large Language Models with Sparse Mixture of Low-Rank Adapter ExpertsCode5
TaskWeaver: A Code-First Agent FrameworkCode5
How to Design Translation Prompts for ChatGPT: An Empirical StudyCode5
Decoder Tuning: Efficient Language Understanding as DecodingCode4
What Makes Good In-Context Examples for GPT-3?Code4
Zero-Shot Learners for Natural Language Understanding via a Unified Multiple Choice PerspectiveCode4
Attention Is All You NeedCode3
Show:102550

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1ConvBERT-DGAverage74.6—Unverified
2ConvBERT-DG + Pre + MultiAverage73.8—Unverified
3mslmAverage73.49—Unverified
4ConvBERT + Pre + MultiAverage68.22—Unverified
5BanLanGenAverage39.16—Unverified
#ModelMetricClaimedVerifiedStatus
1ConvBERT + Pre + MultiAverage86.89—Unverified
2mslmAverage85.83—Unverified
3ConvBERT-DG + Pre + MultiAverage85.34—Unverified
#ModelMetricClaimedVerifiedStatus
1Subword-level Transformer LMAccuracy58.3—Unverified
#ModelMetricClaimedVerifiedStatus
1RoBERTa + LinearFull F1 (Preps)78.2—Unverified