| WinoWhat: A Parallel Corpus of Paraphrased WinoGrande Sentences with Common Sense Categorization | Mar 31, 2025 | Common Sense ReasoningMemorization | —Unverified | 0 |
| Obliviate: Efficient Unmemorization for Protecting Intellectual Property in Large Language Models | Feb 20, 2025 | HellaSwagMemorization | —Unverified | 0 |
| PortLLM: Personalizing Evolving Large Language Models with Training-Free and Portable Model Patches | Oct 8, 2024 | GPUGSM8K | —Unverified | 0 |
| Judgment of Thoughts: Courtroom of the Binary Logical Reasoning in Large Language Models | Sep 25, 2024 | Fake News DetectionLanguage Modeling | —Unverified | 0 |
| metabench -- A Sparse Benchmark to Measure General Ability in Large Language Models | Jul 4, 2024 | ARCGSM8K | CodeCode Available | 0 |
| Promises, Outlooks and Challenges of Diffusion Language Modeling | Jun 17, 2024 | ARCHellaSwag | —Unverified | 0 |
| Who's Harry Potter? Approximate Unlearning in LLMs | Oct 3, 2023 | ARCGPU | —Unverified | 0 |
| Are Hard Examples also Harder to Explain? A Study with Human and Model-Generated Explanations | Nov 14, 2022 | Winogrande | CodeCode Available | 0 |
| On Curriculum Learning for Commonsense Reasoning | Jul 1, 2022 | HellaSwagLearning-To-Rank | CodeCode Available | 0 |
| A Warm Start and a Clean Crawled Corpus - A Recipe for Good Language Models | Jun 1, 2022 | Constituency ParsingGrammatical Error Detection | —Unverified | 0 |