SOTAVerified

Image to text

Papers

Showing 176–200 of 246 papers

TitleStatusHype
Dallah: A Dialect-Aware Multimodal Large Language Model for Arabic—0
DART: Disease-aware Image-Text Alignment and Self-correcting Re-alignment for Trustworthy Radiology Report Generation—0
Deductron -- A Recurrent Neural Network—0
Development of a New Image-to-text Conversion System for Pashto, Farsi and Traditional Chinese—0
DiffusionSTR: Diffusion Model for Scene Text Recognition—0
DiffuVST: Narrating Fictional Scenes with Global-History-Guided Denoising Models—0
DIR: Retrieval-Augmented Image Captioning with Comprehensive Understanding—0
Discovering Bugs in Vision Models using Off-the-shelf Image Generation and Captioning—0
Doc2Im: document to image conversion through self-attentive embedding—0
DOCCI: Descriptions of Connected and Contrasting Images—0
Do DALL-E and Flamingo Understand Each Other?—0
Do LLMs Understand Visual Anomalies? Uncovering LLM's Capabilities in Zero-shot Anomaly Detection—0
Dynamic Traceback Learning for Medical Report Generation—0
Efficient End-to-End Visual Document Understanding with Rationale Distillation—0
EI-CLIP: Entity-Aware Interventional Contrastive Learning for E-Commerce Cross-Modal Retrieval—0
EmojiGAN: learning emojis distributions with a generative model—0
Enhancing Vision-Language Pre-training with Rich Supervisions—0
Evaluating authenticity and quality of image captions via sentiment and semantic analyses—0
Every picture tells a story: Image-grounded controllable stylistic story generation—0
Everything is a Video: Unifying Modalities through Next-Frame Prediction—0
Eyes Closed, Safety On: Protecting Multimodal LLMs via Image-to-Text Transformation—0
Faithful Chart Summarization with ChaTS-Pi—0
Fetch-A-Set: A Large-Scale OCR-Free Benchmark for Historical Document Retrieval—0
From Image to Text Classification: A Novel Approach based on Clustering Word Embeddings—0
From Image to Text in Sentiment Analysis via Regression and Deep Learning—0
Show:102550
← PrevPage 8 of 10Next →

No leaderboard results yet.