| Benchmarking a Benchmark: How Reliable is MS-COCO? | Nov 5, 2023 | Benchmarkingimage-classification | —Unverified | 0 | 0 |
| PASTA: A Dataset for Modeling Participant States in Narratives | Jul 31, 2022 | BenchmarkingCommon Sense Reasoning | —Unverified | 0 | 0 |
| Yambda-5B -- A Large-Scale Multi-modal Dataset for Ranking And Retrieval | May 28, 2025 | BenchmarkingRecommendation Systems | —Unverified | 0 | 0 |
| PatentNet: A Large-Scale Incomplete Multiview, Multimodal, Multilabel Industrial Goods Image Database | Jun 23, 2021 | BenchmarkingClustering | —Unverified | 0 | 0 |
| PathBench: A Benchmarking Platform for Classical and Learned Path Planning Algorithms | May 4, 2021 | Benchmarking | —Unverified | 0 | 0 |
| PathBench: A comprehensive comparison benchmark for pathology foundation models towards precision oncology | May 26, 2025 | BenchmarkingPrognosis | —Unverified | 0 | 0 |
| Patherea: Cell Detection and Classification for the 2020s | Dec 21, 2024 | BenchmarkingCell Detection | —Unverified | 0 | 0 |
| A Correlation- and Mean-Aware Loss Function and Benchmarking Framework to Improve GAN-based Tabular Data Synthesis | May 27, 2024 | Benchmarking | —Unverified | 0 | 0 |
| A Continuously Growing Dataset of Sentential Paraphrases | Aug 1, 2017 | BenchmarkingParaphrase Identification | —Unverified | 0 | 0 |
| Pathway: a fast and flexible unified stream data processing framework for analytical and Machine Learning applications | Jul 12, 2023 | Benchmarking | —Unverified | 0 | 0 |