SOTAVerified

Large Language Model

Papers

Showing 801–850 of 6097 papers

TitleStatusHype
System of Agentic AI for the Discovery of Metal-Organic Frameworks—0
PV-VLM: A Multimodal Vision-Language Approach Incorporating Sky Images for Intra-Hour Photovoltaic Power Forecasting—0
Chain-of-Thought Textual Reasoning for Few-shot Temporal Action Localization—0
Towards a Multi-Agent Vision-Language System for Zero-Shot Novel Hazardous Object Detection for Autonomous Driving SafetyCode0
Zero-Shot Industrial Anomaly Segmentation with Image-Aware Prompt GenerationCode0
Scaling sparse feature circuit finding for in-context learning—0
RAG Without the Lag: Interactive Debugging for Retrieval-Augmented Generation Pipelines—0
Are Retrials All You Need? Enhancing Large Language Model Reasoning Without Verbalized Feedback—0
Causal-Copilot: An Autonomous Causal Analysis Agent—0
ChatEXAONEPath: An Expert-level Multimodal Large Language Model for Histopathology Using Whole Slide Images—0
Pandora: A Code-Driven Large Language Model Agent for Unified Reasoning Across Diverse Structured Knowledge—0
Enhancing the Geometric Problem-Solving Ability of Multimodal LLMs via Symbolic-Neural IntegrationCode1
Can LLMs reason over extended multilingual contexts? Towards long-context evaluation beyond retrieval and haystacksCode0
Retrieval-Augmented Generation with Conflicting EvidenceCode1
DIDS: Domain Impact-aware Data Sampling for Large Language Model Training—0
EarthGPT-X: Enabling MLLMs to Flexibly and Comprehensively Understand Multi-Source Remote Sensing Imagery—0
SkyReels-V2: Infinite-length Film Generative ModelCode9
Uncertainty-Aware Trajectory Prediction via Rule-Regularized Heteroscedastic Deep ClassificationCode0
SmartFreeEdit: Mask-Free Spatial-Aware Image Editing with Complex Instruction UnderstandingCode1
Mixer Metaphors: audio interfaces for non-musical applications—0
BitNet b1.58 2B4T Technical Report—0
Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM—0
Trusting CHATGPT: how minor tweaks in the prompts lead to major differences in sentiment classification—0
AnomalyR1: A GRPO-based End-to-end MLLM for Industrial Anomaly DetectionCode1
Generative Recommendation with Continuous-Token Diffusion—0
Rethinking LLM-Based Recommendations: A Query Generation-Based, Training-Free Approach—0
HLS-Eval: A Benchmark and Framework for Evaluating LLMs on High-Level Synthesis Design TasksCode1
Position: The Most Expensive Part of an LLM should be its Training Data—0
Characterizing and Optimizing LLM Inference Workloads on CPU-GPU Coupled Architectures—0
Towards Conversational AI for Human-Machine Collaborative MLOps—0
Recommending Clinical Trials for Online Patient Cases using Artificial Intelligence—0
GraphicBench: A Planning Benchmark for Graphic Design with Language Agents—0
A Large-Language Model Framework for Relative Timeline Extraction from PubMed Case Reports—0
Video Summarization with Large Language Models—0
When is Task Vector Provably Effective for Model Editing? A Generalization Analysis of Nonlinear Transformers—0
Large Language Model-Informed Feature Discovery Improves Prediction and Interpretation of Credibility Perceptions of Visual Content—0
ReZero: Enhancing LLM search ability by trying one-more-time—0
Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement LearningCode3
Learning to Be A Doctor: Searching for Effective Medical Agent Architectures—0
The Obvious Invisible Threat: LLM-Powered GUI Agents' Vulnerability to Fine-Print Injections—0
Evaluation Report on MCP ServersCode3
Transferable text data distillation by trajectory matching—0
A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science—0
Investigating cybersecurity incidents using large language models in latest-generation wireless networks—0
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current BenchmarksCode0
Mavors: Multi-granularity Video Representation for Multimodal Large Language Model—0
The Scalability of Simplicity: Empirical Analysis of Vision-Language Learning with a Single TransformerCode2
InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models—0
LangPert: Detecting and Handling Task-level Perturbations for Robust Object Rearrangement—0
Automated Testing of COBOL to Java Transformation—0
Show:102550
← PrevPage 17 of 122Next →

No leaderboard results yet.