SOTAVerified

Zero-shot Generalization

Papers

Showing 501–550 of 572 papers

TitleStatusHype
Scoring-Aggregating-Planning: Learning task-agnostic priors from interactions and sparse rewards for zero-shot generalization—0
Segment Anything Model for Grain Characterization in Hard Drive Design—0
Select and Distill: Selective Dual-Teacher Knowledge Transfer for Continual Learning on Vision-Language Models—0
Self-FiLM: Conditioning GANs with self-supervised representations for bandwidth extension based speaker recognition—0
Hint of Thought prompting: an explainable and zero-shot approach to reasoning tasks with LLMs—0
Sequence-Based Plan Feasibility Prediction for Efficient Task and Motion Planning—0
Show, Don’t Tell: Demonstrations Outperform Descriptions for Schema-Guided Task-Oriented Dialogue—0
Show, Don't Tell: Demonstrations Outperform Descriptions for Schema-Guided Task-Oriented Dialogue—0
SimSort: A Data-Driven Framework for Spike Sorting by Large-Scale Electrophysiology Simulation—0
Solving Continual Offline Reinforcement Learning with Decision Transformer—0
Solving the Same-Different Task with Convolutional Neural Networks—0
SPT: Semi-Parametric Prompt Tuning for Multitask Prompted Learning—0
SSTD: Stripe-Like Space Target Detection Using Single-Point Weak Supervision—0
State Combinatorial Generalization In Decision Making With Conditional Diffusion Models—0
StereoGen: High-quality Stereo Image Generation from a Single Image—0
Still not systematic after all these years: On the compositional skills of sequence-to-sequence recurrent networks—0
Style-Pro: Style-Guided Prompt Learning for Generalizable Vision-Language Models—0
StyLIP: Multi-Scale Style-Conditioned Prompt Learning for CLIP-based Domain Generalization—0
Survey on Monocular Metric Depth Estimation—0
TanDepth: Leveraging Global DEMs for Metric Monocular Depth Estimation in UAVs—0
A Dual Curriculum Learning Framework for Multi-UAV Pursuit-Evasion in Diverse Environments—0
ConfusionPrompt: Practical Private Inference for Online Large Language Models—0
Test-time Loss Landscape Adaptation for Zero-Shot Generalization in Vision-Language Models—0
Text2Model: Text-based Model Induction for Zero-shot Image Classification—0
Text-only Synthesis for Image Captioning—0
Text-to-Decision Agent: Learning Generalist Policies from Natural Language Supervision—0
The Matrix: Infinite-Horizon World Generation with Real-Time Moving Control—0
The Third Monocular Depth Estimation Challenge—0
Thinking agents for zero-shot generalization to qualitatively novel tasks—0
TIMA: Text-Image Mutual Awareness for Balancing Zero-Shot Adversarial Robustness and Generalization Ability—0
TimeGraphs: Graph-based Temporal Reasoning—0
Towards Artificial General or Personalized Intelligence? A Survey on Foundation Models for Personalized Federated Intelligence—0
Towards Depth Foundation Model: Recent Trends in Vision-Based Depth Estimation—0
Towards Generalist Biomedical AI—0
Towards the Unification of Generative and Discriminative Visual Foundation Model: A Survey—0
Towards Vision-Language-Garment Models For Web Knowledge Garment Understanding and Generation—0
Toward Task Generalization via Memory Augmentation in Meta-Reinforcement Learning—0
Transductive CLIP with Class-Conditional Contrastive Learning—0
Transferable and Distributed User Association Policies for 5G and Beyond Networks—0
Quantifying uncertainty in lung cancer segmentation with foundation models applied to mixed-domain datasets—0
Turbocharging Solution Concepts: Solving NEs, CEs and CCEs with Neural Equilibrium Solvers—0
Unifying Few- and Zero-Shot Egocentric Action Recognition—0
Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos—0
UniIR: Training and Benchmarking Universal Multimodal Information Retrievers—0
Unpaired Object-Level SAR-to-Optical Image Translation for Aircraft with Keypoints-Guided Diffusion Models—0
Unsupervised Discovery of Object-Centric Neural Fields—0
Unsupervised Prompt Tuning for Text-Driven Object Detection—0
UTSD: Unified Time Series Diffusion Model—0
Value Function Spaces: Skill-Centric State Abstractions for Long-Horizon Reasoning—0
Video Event Reasoning and Prediction by Fusing World Knowledge from LLMs with Vision Foundation Models—0
Show:102550
← PrevPage 11 of 12Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1GR-MGAvg. sequence length4.04—Unverified
2MoDEAvg. sequence length4.01—Unverified
3RoboUniViewAvg. sequence length3.65—Unverified
43D Diffuser ActorAvg. sequence length3.27—Unverified
5GR-1Avg. sequence length3.06—Unverified