Question Answering

Question answering can be segmented into domain-specific tasks like community question answering and knowledge-base question answering. Popular benchmark datasets for evaluation question answering systems include SQuAD, HotPotQA, bAbI, TriviaQA, WikiQA, and many others. Models for question answering are typically evaluated on metrics like EM and F1. Some recent top performing models are T5 and XLNet.

( Image credit: SQuAD )

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 7926–7950 of 10817 papers

Title	Date	Tasks	Status
Raccoons at SemEval-2022 Task 11: Leveraging Concatenated Word Embeddings for Named Entity Recognition	Jul 1, 2022	Machine Translationnamed-entity-recognition	—Unverified
Complex Question Answering: Unsupervised Learning Approaches and Experiments	Jan 15, 2014	Document SummarizationMulti-Document Summarization	—Unverified
RA-DIT: Retrieval-Augmented Dual Instruction Tuning	Oct 2, 2023	Few-Shot LearningOpen-Domain Question Answering	—Unverified
RadQA: A Question Answering Dataset to Improve Comprehension of Radiology Reports	Jun 1, 2022	Question AnsweringReading Comprehension	—Unverified
AssistPDA: An Online Video Surveillance Assistant for Video Anomaly Prediction, Detection, and Analysis	Mar 27, 2025	Anomaly DetectionAnomaly Forecasting	—Unverified
Assistive Image Annotation Systems with Deep Learning and Natural Language Capabilities: A Review	Jun 28, 2024	Active LearningImage Captioning	—Unverified
RAG based Question-Answering for Contextual Response Prediction System	Sep 5, 2024	PredictionQuestion Answering	—Unverified
RAG-based Question Answering over Heterogeneous Data and Text	Dec 10, 2024	Answer GenerationKnowledge Graphs	—Unverified
Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models	Feb 3, 2025	Question AnsweringRAG	—Unverified
GoT-CQA: Graph-of-Thought Guided Compositional Reasoning for Chart Question Answering	Sep 4, 2024	Chart Question AnsweringQuestion Answering	—Unverified
Goodwill Hunting: Analyzing and Repurposing Off-the-Shelf Named Entity Linking Systems	Jun 1, 2021	Entity LinkingQuestion Answering	—Unverified
Complex Question Answering on knowledge graphs using machine translation and multi-task learning	Apr 1, 2021	Entity LinkingKnowledge Graphs	—Unverified
Good, Great, Excellent: Global Inference of Semantic Intensities	Jan 1, 2013	Natural Language InferenceQuestion Answering	—Unverified
RAG-RL: Advancing Retrieval-Augmented Generation via RL and Curriculum Learning	Mar 17, 2025	Answer GenerationMulti-hop Question Answering	—Unverified
FOCUS: Internal MLLM Representations for Efficient Fine-Grained Visual Question Answering	Jun 25, 2025	Question AnsweringVisual Question Answering	—Unverified
RAG vs. GraphRAG: A Systematic Evaluation and Key Insights	Feb 17, 2025	Knowledge GraphsQuestion Answering	—Unverified
Examining Long-Context Large Language Models for Environmental Review Document Comprehension	Jul 10, 2024	Question AnsweringRAG	—Unverified
Complex QA and language models hybrid architectures, Survey	Feb 17, 2023	Domain AdaptationFairness	—Unverified
Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts	Feb 26, 2024	DiversityQuestion Answering	—Unverified
Reasoning and Tools for Human-Level Forecasting	Aug 21, 2024	Decision MakingQuestion Answering	—Unverified
Good, Better, Best: Textual Distractors Generation for Multiple-Choice Visual Question Answering via Reinforcement Learning	Oct 21, 2019	Data AugmentationDecision Making	—Unverified
Complex Program Induction for Querying Knowledge Bases in the Absence of Gold Programs	Mar 1, 2019	Natural Language QueriesProgram induction	—Unverified
GOF at Qur’an QA 2022: Towards an Efficient Question Answering For The Holy Qu’ran In The Arabic Language Using Deep Learning-Based Approach	Jun 1, 2022	Question Answering	—Unverified
FONDUE: A Framework for Node Disambiguation Using Network Embeddings	Feb 24, 2020	Knowledge GraphsNetwork Embedding	—Unverified
Assisting Scene Graph Generation with Self-Supervision	Aug 8, 2020	Graph GenerationImage Captioning	—Unverified

Show:10 25 50

← PrevPage 318 of 433Next →

All datasets SQuAD2.0 SQuAD1.1 HotpotQA PIQA BoolQ COPA TriviaQA SQuAD1.1 dev Natural Questions OpenBookQA TruthfulQA MultiRC

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	IE-Net (ensemble)	EM	90.94	—	Unverified
2	FPNet (ensemble)	EM	90.87	—	Unverified
3	IE-NetV2 (ensemble)	EM	90.86	—	Unverified
4	SA-Net on Albert (ensemble)	EM	90.72	—	Unverified
5	SA-Net-V2 (ensemble)	EM	90.68	—	Unverified
6	FPNet (ensemble)	EM	90.6	—	Unverified
7	Retro-Reader (ensemble)	EM	90.58	—	Unverified
8	EntitySpanFocusV2 (ensemble)	EM	90.52	—	Unverified
9	TransNets + SFVerifier + SFEnsembler (ensemble)	EM	90.49	—	Unverified
10	EntitySpanFocus+AT (ensemble)	EM	90.45	—	Unverified