SOTAVerified

World Knowledge

Papers

Showing 51–100 of 818 papers

TitleStatusHype
EventVAD: Training-Free Event-Aware Video Anomaly Detection—0
Can GPT tell us why these images are synthesized? Empowering Multimodal Large Language Models for Forensics—0
Rethinking LLM-Based Recommendations: A Query Generation-Based, Training-Free Approach—0
Coding-Prior Guided Diffusion Network for Video Deblurring—0
ARise: Towards Knowledge-Augmented Reasoning via Risk-Adaptive Search—0
Enhancing LLM-based Recommendation through Semantic-Aligned Collaborative Knowledge—0
Large Language Model Empowered Recommendation Meets All-domain Continual Pre-Training—0
ConceptFormer: Towards Efficient Use of Knowledge-Graph Embeddings in Large Language Models—0
DiffusionCom: Structure-Aware Multimodal Diffusion Model for Multimodal Knowledge Graph Completion—0
Have we unified image generation and understanding yet? An empirical study of GPT-4o's image generation ability—0
Large Language Models Enhanced Hyperbolic Space Recommender Systems—0
Memory-Modular Classification: Learning to Generalize with Memory ReplacementCode0
User Feedback Alignment for LLM-powered Exploration in Large-scale Recommendation Systems—0
RS-RAG: Bridging Remote Sensing Imagery and Comprehensive Knowledge with a Multi-Modal Dataset and Retrieval-Augmented Generation Model—0
Adaptive Elicitation of Latent Information Using Natural Language—0
Knowledge Graph Completion with Mixed Geometry Tensor FactorizationCode0
GPT-ImgEval: A Comprehensive Benchmark for Diagnosing GPT4o in Image GenerationCode3
F-ViTA: Foundation Model Guided Visible to Thermal TranslationCode1
OnRL-RAG: Real-Time Personalized Mental Health Dialogue System—0
A Diffusion-Based Framework for Occluded Object Movement—0
Generative Retrieval and Alignment Model: A New Paradigm for E-commerce Retrieval—0
Synthetic-to-Real Self-supervised Robust Depth Estimation via Learning with Motion and Structure PriorsCode1
LLM-based Agent Simulation for Maternal Health Interventions: Uncertainty Estimation and Decision-focused EvaluationCode0
Test-Time Reasoning Through Visual Human Preferences with VLMs and Soft Rewards—0
Human-Object Interaction with Vision-Language Model Guided Relative Movement Dynamics—0
Instructing the Architecture Search for Spatial-temporal Sequence Forecasting with LLM—0
A Study into Investigating Temporal Robustness of LLMs—0
Advancing Problem-Based Learning in Biomedical Engineering in the Era of Generative AI—0
World Knowledge from AI Image Generation for Robot Control—0
JARVIS-VLA: Post-Training Large-Scale Vision Language Models to Play Visual Games with Keyboards and Mouse—0
Exploiting Diffusion Prior for Real-World Image Dehazing with Unpaired TrainingCode1
FusDreamer: Label-efficient Remote Sensing World Model for Multimodal Data ClassificationCode1
Impossible Videos—0
A Multi-Stage Framework with Taxonomy-Guided Reasoning for Occupation Classification Using Large Language Models—0
Free-form language-based robotic reasoning and graspingCode2
A Framework for a Capability-driven Evaluation of Scenario Understanding for Multimodal Large Language Models in Autonomous Driving—0
Who Relies More on World Knowledge and Bias for Syntactic Ambiguity Resolution: Humans or LLMs?Code0
SySLLM: Generating Synthesized Policy Summaries for Reinforcement Learning Agents Using Large Language Models—0
LREF: A Novel LLM-based Relevance Framework for E-commerce—0
WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image GenerationCode4
PointVLA: Injecting the 3D World into Vision-Language-Action ModelsCode4
The Society of HiveMind: Multi-Agent Optimization of Foundation Model Swarms to Unlock the Potential of Collective Intelligence—0
Effective LLM Knowledge Learning via Model Generalization—0
From Language to Cognition: How LLMs Outgrow the Human Language Network—0
Can Large Language Models Help Experimental Design for Causal Discovery?—0
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds—0
FaithUn: Toward Faithful Forgetting in Language Models by Investigating the Interconnectedness of Knowledge—0
Data-Efficient Multi-Agent Spatial Planning with LLMs—0
BottleHumor: Self-Informed Humor Explanation using the Information Bottleneck PrincipleCode0
LLM4Tag: Automatic Tagging System for Information Retrieval via Large Language Models—0
Show:102550
← PrevPage 2 of 17Next →

No leaderboard results yet.