| Generating Out-Of-Distribution Scenarios Using Language Models | Nov 25, 2024 | Autonomous DrivingAutonomous Vehicles | —Unverified | 0 |
| Context-Aware Multimodal Pretraining | Nov 22, 2024 | Contrastive LearningRepresentation Learning | —Unverified | 0 |
| HEIGHT: Heterogeneous Interaction Graph Transformer for Robot Navigation in Crowded and Constrained Environments | Nov 19, 2024 | Deep Reinforcement LearningRobot Navigation | —Unverified | 0 |
| SAM Carries the Burden: A Semi-Supervised Approach Refining Pseudo Labels for Medical Segmentation | Nov 19, 2024 | Image SegmentationMedical Image Segmentation | CodeCode Available | 0 |
| Scalable Autoregressive Monocular Depth Estimation | Nov 18, 2024 | Depth EstimationMonocular Depth Estimation | —Unverified | 0 |
| MLAN: Language-Based Instruction Tuning Improves Zero-Shot Generalization of Multimodal Large Language Models | Nov 15, 2024 | Instruction FollowingZero-shot Generalization | CodeCode Available | 0 |
| Self-Supervised Monocular 4D Scene Reconstruction for Egocentric Videos | Nov 14, 2024 | 4D reconstructionSelf-Supervised Learning | —Unverified | 0 |
| Mono2Stereo: Monocular Knowledge Transfer for Enhanced Stereo Matching | Nov 14, 2024 | Depth EstimationKnowledge Distillation | —Unverified | 0 |
| WorkflowLLM: Enhancing Workflow Orchestration Capability of Large Language Models | Nov 8, 2024 | Task PlanningZero-shot Generalization | CodeCode Available | 2 |
| In the Era of Prompt Learning with Vision-Language Models | Nov 7, 2024 | Domain AdaptationDomain Generalization | —Unverified | 0 |