Question Answering

Question answering can be segmented into domain-specific tasks like community question answering and knowledge-base question answering. Popular benchmark datasets for evaluation question answering systems include SQuAD, HotPotQA, bAbI, TriviaQA, WikiQA, and many others. Models for question answering are typically evaluated on metrics like EM and F1. Some recent top performing models are T5 and XLNet.

( Image credit: SQuAD )

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 1501–1550 of 10817 papers

Title	Date	Tasks	Status	Hype	Score
Collab-RAG: Boosting Retrieval-Augmented Generation for Complex Question Answering via White-Box and Black-Box LLM Collaboration	Apr 7, 2025	Language ModelingLanguage Modelling	CodeCode Available	1	5
Eliminating Position Bias of Language Models: A Mechanistic Approach	Jul 1, 2024	Mathobject-detection	CodeCode Available	1	5
COBRA: Contrastive Bi-Modal Representation Algorithm	May 7, 2020	Cross-Modal RetrievalImage Captioning	CodeCode Available	1	5
IntentionQA: A Benchmark for Evaluating Purchase Intention Comprehension Abilities of Language Models in E-commerce	Jun 14, 2024	Multiple-choiceQuestion Answering	CodeCode Available	1	5
BiMediX: Bilingual Medical Mixture of Experts LLM	Feb 20, 2024	Mixture-of-ExpertsMultiple-choice	CodeCode Available	1	5
Emergence of Grounded Compositional Language in Multi-Agent Populations	Mar 15, 2017	Machine TranslationQuestion Answering	CodeCode Available	1	5
Interconnected Question Generation with Coreference Alignment and Conversation Flow Modeling	Jun 17, 2019	Question AnsweringQuestion Generation	CodeCode Available	1	5
MISS: A Generative Pretraining and Finetuning Approach for Med-VQA	Jan 10, 2024	Medical Visual Question AnsweringMulti-Task Learning	CodeCode Available	1	5
BioBERT: a pre-trained biomedical language representation model for biomedical text mining	Jan 25, 2019	Drug–drug Interaction ExtractionFew-Shot Learning	CodeCode Available	1	5
BioBridge: Bridging Biomedical Foundation Models via Knowledge Graphs	Oct 5, 2023	Cross-Modal RetrievalDomain Generalization	CodeCode Available	1	5
IoT-LM: Large Multisensory Language Models for the Internet of Things	Jul 13, 2024	Language ModelingLanguage Modelling	CodeCode Available	1	5
BioELECTRA:Pretrained Biomedical text Encoder using Discriminators	Jun 11, 2021	ArticlesLanguage Modeling	CodeCode Available	1	5
KGE-CL: Contrastive Learning of Tensor Decomposition Based Knowledge Graph Embeddings	Dec 9, 2021	Contrastive LearningGraph Embedding	CodeCode Available	1	5
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models	Feb 17, 2024	Medical Question AnsweringQuestion Answering	CodeCode Available	1	5
emrQA: A Large Corpus for Question Answering on Electronic Medical Records	Sep 3, 2018	FormQuestion Answering	CodeCode Available	1	5
Mitigating the Position Bias of Transformer Models in Passage Re-Ranking	Jan 18, 2021	Passage Re-RankingPosition	CodeCode Available	1	5
Encoding and Controlling Global Semantics for Long-form Video Question Answering	May 30, 2024	FormQuestion Answering	CodeCode Available	1	5
Learning Associative Inference Using Fast Weight Memory	Nov 16, 2020	Language ModellingMeta Reinforcement Learning	CodeCode Available	1	5
A Systematic Study and Comprehensive Evaluation of ChatGPT on Benchmark Datasets	May 29, 2023	Bias DetectionCode Generation	CodeCode Available	1	5
-Video: A Training-Free Approach to Long Video Understanding via Continuous-Time Memory Consolidation	Jan 31, 2025	Question AnsweringVideo Question Answering	CodeCode Available	1	5
Infusing Disease Knowledge into BERT for Health Question Answering, Medical Inference and Disease Name Recognition	Oct 8, 2020	Question AnsweringWorld Knowledge	CodeCode Available	1	5
Enhancing Multi-modal and Multi-hop Question Answering via Structured Knowledge and Unified Retrieval-Generation	Dec 16, 2022	Answer GenerationDecoder	CodeCode Available	1	5
End-to-End Training of Multi-Document Reader and Retriever for Open-Domain Question Answering	Jun 9, 2021	Answer GenerationOpen-Domain Question Answering	CodeCode Available	1	5
Engineering flexible machine learning systems by traversing functionally-invariant paths	Apr 30, 2022	Adversarial RobustnessContinual Learning	CodeCode Available	1	5
Clues Before Answers: Generation-Enhanced Multiple-Choice QA	Apr 30, 2022	DecoderMultiple-choice	CodeCode Available	1	5
End-to-End Training of Neural Retrievers for Open-Domain Question Answering	Jan 2, 2021	Natural QuestionsOpen-Domain Question Answering	CodeCode Available	1	5
A Symmetric Dual Encoding Dense Retrieval Framework for Knowledge-Intensive Visual Question Answering	Apr 26, 2023	DecoderKnowledge Distillation	CodeCode Available	1	5
MMBERT: Multimodal BERT Pretraining for Improved Medical VQA	Apr 3, 2021	Language ModelingLanguage Modelling	CodeCode Available	1	5
Enhancing Contextual Understanding in Large Language Models through Contrastive Decoding	May 4, 2024	Open-Domain Question AnsweringQuestion Answering	CodeCode Available	1	5
Information Theoretic Representation Distillation	Dec 1, 2021	Classification with Binary Weight NetworkKnowledge Distillation	CodeCode Available	1	5
Enhancing In-Context Learning with Answer Feedback for Multi-Span Question Answering	Jun 7, 2023	In-Context LearningKeyphrase Extraction	CodeCode Available	1	5
Enhancing LLM's Cognition via Structurization	Jul 23, 2024	HallucinationHallucination Evaluation	CodeCode Available	1	5
Initial Nugget Evaluation Results for the TREC 2024 RAG Track with the AutoNuggetizer Framework	Nov 14, 2024	Question AnsweringRAG	CodeCode Available	1	5
CL-ReLKT: Cross-lingual Language Knowledge Transfer for Multilingual Retrieval Question Answering	Jul 1, 2022	Language ModelingLanguage Modelling	CodeCode Available	1	5
SentenceMIM: A Latent Variable Language Model	Feb 18, 2020	Language ModelingLanguage Modelling	CodeCode Available	1	5
CLTR: An End-to-End, Transformer-Based System for Cell Level Table Retrieval and Table Question Answering	Jun 8, 2021	Question AnsweringRetrieval	CodeCode Available	1	5
Enhancing Vision-Language Pre-Training with Jointly Learned Questioner and Dense Captioner	May 19, 2023	Dense CaptioningImage Captioning	CodeCode Available	1	5
ERICA: Improving Entity and Relation Understanding for Pre-trained Language Models via Contrastive Learning	Dec 30, 2020	Contrastive LearningEntity Typing	CodeCode Available	1	5
BioPlanner: Automatic Evaluation of LLMs on Protocol Planning in Biology	Oct 16, 2023	Language ModellingQuestion Answering	CodeCode Available	1	5
BioProBench: Comprehensive Dataset and Benchmark in Biological Protocol Understanding and Reasoning	May 11, 2025	Question Answering	CodeCode Available	1	5
Modeling Worlds in Text	May 21, 2021	Action ParsingKnowledge Graphs	CodeCode Available	1	5
Model Internals-based Answer Attribution for Trustworthy Retrieval-Augmented Generation	Jun 19, 2024	Question AnsweringRAG	CodeCode Available	1	5
Clover: Towards A Unified Video-Language Alignment and Fusion Model	Jul 16, 2022	Language ModelingLanguage Modelling	CodeCode Available	1	5
InfoBERT: Improving Robustness of Language Models from An Information Theoretic Perspective	Oct 5, 2020	Natural Language InferenceQuestion Answering	CodeCode Available	1	5
InforMask: Unsupervised Informative Masking for Language Model Pretraining	Oct 21, 2022	Language ModelingLanguage Modelling	CodeCode Available	1	5
EntQA: Entity Linking as Question Answering	Oct 5, 2021	BenchmarkingEntity Linking	CodeCode Available	1	5
Entailment Tree Explanations via Iterative Retrieval-Generation Reasoner	May 18, 2022	DecoderQuestion Answering	CodeCode Available	1	5
Entity-Based Knowledge Conflicts in Question Answering	Sep 10, 2021	HallucinationOut-of-Distribution Generalization	CodeCode Available	1	5
Answering Ambiguous Questions through Generative Evidence Fusion and Round-Trip Prediction	Nov 26, 2020	Open-Domain Question AnsweringQuestion Answering	CodeCode Available	1	5
Injecting Numerical Reasoning Skills into Language Models	Apr 9, 2020	Data AugmentationDecoder	CodeCode Available	1	5

Show:10 25 50

← PrevPage 31 of 217Next →

All datasets SQuAD2.0 SQuAD1.1 HotpotQA PIQA BoolQ COPA TriviaQA SQuAD1.1 dev Natural Questions OpenBookQA TruthfulQA MultiRC

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	IE-Net (ensemble)	EM	90.94	—	Unverified
2	FPNet (ensemble)	EM	90.87	—	Unverified
3	IE-NetV2 (ensemble)	EM	90.86	—	Unverified
4	SA-Net on Albert (ensemble)	EM	90.72	—	Unverified
5	SA-Net-V2 (ensemble)	EM	90.68	—	Unverified
6	FPNet (ensemble)	EM	90.6	—	Unverified
7	Retro-Reader (ensemble)	EM	90.58	—	Unverified
8	EntitySpanFocusV2 (ensemble)	EM	90.52	—	Unverified
9	TransNets + SFVerifier + SFEnsembler (ensemble)	EM	90.49	—	Unverified
10	EntitySpanFocus+AT (ensemble)	EM	90.45	—	Unverified