| Change Detection Meets Visual Question Answering | Dec 12, 2021 | Answer GenerationChange Detection | CodeCode Available | 1 | 5 |
| CICERO: A Dataset for Contextualized Commonsense Inference in Dialogues | Mar 25, 2022 | Answer GenerationAnswer Selection | CodeCode Available | 1 | 5 |
| Do RAG Systems Cover What Matters? Evaluating and Optimizing Responses with Sub-Question Coverage | Oct 20, 2024 | Answer GenerationRAG | CodeCode Available | 1 | 5 |
| Answer Mining from a Pool of Images: Towards Retrieval-Based Visual Question Answering | Jun 29, 2023 | Answer GenerationQuestion Answering | CodeCode Available | 1 | 5 |
| Debate on Graph: a Flexible and Reliable Reasoning Framework for Large Language Models | Sep 5, 2024 | Answer GenerationGraph Question Answering | CodeCode Available | 1 | 5 |
| GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents | May 21, 2025 | Answer GenerationReinforcement Learning (RL) | CodeCode Available | 1 | 5 |
| Do Vision & Language Decoders use Images and Text equally? How Self-consistent are their Explanations? | Apr 29, 2024 | Answer GenerationBenchmarking | CodeCode Available | 1 | 5 |
| ECoRAG: Evidentiality-guided Compression for Long Context RAG | Jun 5, 2025 | Answer GenerationOpen-Domain Question Answering | CodeCode Available | 1 | 5 |
| How to think step-by-step: A mechanistic understanding of chain-of-thought reasoning | Feb 28, 2024 | Answer Generation | CodeCode Available | 1 | 5 |
| An Interactive Multi-modal Query Answering System with Retrieval-Augmented Large Language Models | Jul 5, 2024 | Answer GenerationContrastive Learning | CodeCode Available | 1 | 5 |