| Towards Efficient Automatic Self-Pruning of Large Language Models | Feb 20, 2025 | GPU | —Unverified | 0 |
| Distributed U-net model and Image Segmentation for Lung Cancer Detection | Feb 20, 2025 | CPUFederated Learning | —Unverified | 0 |
| Dynamic Low-Rank Sparse Adaptation for Large Language Models | Feb 20, 2025 | CPUGPU | CodeCode Available | 1 |
| TritonBench: Benchmarking Large Language Model Capabilities for Generating Triton Operators | Feb 20, 2025 | BenchmarkingCode Generation | CodeCode Available | 2 |
| Exploring RWKV for Sentence Embeddings: Layer-wise Analysis and Baseline Comparison for Semantic Similarity | Feb 20, 2025 | GPULanguage Modeling | CodeCode Available | 0 |
| Building reliable sim driving agents by scaling self-play | Feb 20, 2025 | Autonomous VehiclesBenchmarking | CodeCode Available | 4 |
| ParallelComp: Parallel Long-Context Compressor for Length Extrapolation | Feb 20, 2025 | 4k8k | —Unverified | 0 |
| Multiscale Byte Language Models -- A Hierarchical Architecture for Causal Million-Length Sequence Modeling | Feb 20, 2025 | DecoderGPU | CodeCode Available | 0 |
| Learning conformational ensembles of proteins based on backbone geometry | Feb 19, 2025 | GPU | —Unverified | 0 |
| FairKV: Balancing Per-Head KV Cache for Fast Multi-GPU Inference | Feb 19, 2025 | GPU | —Unverified | 0 |