| Benchmarking Foundation Models on Exceptional Cases: Dataset Creation and Validation | Oct 23, 2024 | ArticlesBenchmarking | CodeCode Available | 0 |
| CSS: A Large-scale Cross-schema Chinese Text-to-SQL Medical Dataset | May 25, 2023 | BenchmarkingText to SQL | CodeCode Available | 0 |
| Cryo-RALib -- a modular library for accelerating alignment in cryo-EM | Nov 11, 2020 | BenchmarkingGPU | CodeCode Available | 0 |
| What the Weight?! A Unified Framework for Zero-Shot Knowledge Composition | Jan 23, 2024 | Benchmarking | CodeCode Available | 0 |
| STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions | Sep 20, 2024 | BenchmarkingSensitivity | CodeCode Available | 0 |
| Cross-Lingual Text Classification of Transliterated Hindi and Malayalam | Aug 31, 2021 | BenchmarkingClassification | CodeCode Available | 0 |
| Benchmarking Flexible Electric Loads Scheduling Algorithms under Market Price Uncertainty | Feb 4, 2020 | BenchmarkingDecision Making | CodeCode Available | 0 |
| Yum-me: A Personalized Nutrient-based Meal Recommender System | May 25, 2016 | BenchmarkingRecommendation Systems | CodeCode Available | 0 |
| Benchmarking Federated Learning for Semantic Datasets: Federated Scene Graph Generation | Dec 11, 2024 | BenchmarkingFederated Learning | CodeCode Available | 0 |
| Cross-lingual sentiment classification in low-resource Bengali language | Nov 1, 2020 | BenchmarkingClassification | CodeCode Available | 0 |