| Benchmarking Adversarial Robustness of Image Shadow Removal with Shadow-adaptive Attacks | Mar 15, 2024 | Adversarial AttackAdversarial Robustness | —Unverified | 0 | 0 |
| OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents | Jun 19, 2025 | Benchmarking | —Unverified | 0 | 0 |
| oTTC: Object Time-to-Contact for Motion Estimation in Autonomous Driving | May 13, 2024 | AttributeAutonomous Driving | —Unverified | 0 | 0 |
| Benchmarking Adversarial Robustness of Compressed Deep Learning Models | Aug 16, 2023 | Adversarial RobustnessBenchmarking | —Unverified | 0 | 0 |
| Tropical Attention: Neural Algorithmic Reasoning for Combinatorial Algorithms | May 22, 2025 | Adversarial AttackBenchmarking | —Unverified | 0 | 0 |
| Out of Distribution Performance of State of Art Vision Model | Jan 25, 2023 | Benchmarking | —Unverified | 0 | 0 |
| Benchmarking Adversarial Robustness | Dec 26, 2019 | Adversarial AttackAdversarial Robustness | —Unverified | 0 | 0 |
| Overconfident Oracles: Limitations of In Silico Sequence Design Benchmarking | Feb 24, 2025 | Benchmarking | —Unverified | 0 | 0 |
| Overview and practical recommendations on using Shapley Values for identifying predictive biomarkers via CATE modeling | May 2, 2025 | Benchmarking | —Unverified | 0 | 0 |
| Overview of Todai Robot Project and Evaluation Framework of its NLP-based Problem Solving | May 1, 2014 | Benchmarking | —Unverified | 0 | 0 |