| Foundation Models Knowledge Distillation For Battery Capacity Degradation Forecast | May 13, 2025 | Knowledge DistillationTime Series | CodeCode Available | 1 |
| Towards Artificial General or Personalized Intelligence? A Survey on Foundation Models for Personalized Federated Intelligence | May 11, 2025 | Computational EfficiencyFederated Learning | —Unverified | 0 |
| Learning Graph Representation of Agent Diffusers | May 10, 2025 | Graph Neural NetworkImage Generation | CodeCode Available | 0 |
| Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization | May 8, 2025 | Object LocalizationWeakly-Supervised Object Localization | —Unverified | 0 |
| Benchmarking Vision, Language, & Action Models in Procedurally Generated, Open Ended Action Environments | May 8, 2025 | BenchmarkingPrompt Engineering | CodeCode Available | 1 |
| TeDA: Boosting Vision-Lanuage Models for Zero-Shot 3D Object Retrieval via Testing-time Distribution Alignment | May 5, 2025 | 3D Object RetrievalLanguage Modeling | CodeCode Available | 0 |
| Towards Ball Spin and Trajectory Analysis in Table Tennis Broadcast Videos via Physically Grounded Synthetic-to-Real Transfer | Apr 28, 2025 | Monocular 3D Object LocalizationSports Analytics | CodeCode Available | 1 |
| A Review of 3D Object Detection with Vision-Language Models | Apr 25, 2025 | 3D Object DetectionObject | —Unverified | 0 |
| Text-to-Decision Agent: Learning Generalist Policies from Natural Language Supervision | Apr 21, 2025 | MuJoCoZero-shot Generalization | —Unverified | 0 |
| Dysarthria Normalization via Local Lie Group Transformations for Robust ASR | Apr 16, 2025 | Robust Speech Recognitionspeech-recognition | CodeCode Available | 0 |