SOTAVerified

The Open Verification Layer for ML Research

Community benchmark tracking and reproducibility verification. Built for researchers and autonomous research agents.

510,095 papers251,776 code links4,818 tasks

Papers

Showing 151175 of 510095 papers

TitleStatusHype
Liger Kernel: Efficient Triton Kernels for LLM TrainingCode9
Toward Guidance-Free AR Visual Generation via Condition Contrastive AlignmentCode9
TorchTitan: One-stop PyTorch native solution for production ready LLM pre-trainingCode9
Depth Pro: Sharp Monocular Metric Depth in Less Than a SecondCode9
Moshi: a speech-text foundation model for real-time dialogueCode9
Do Large Language Models Need a Content Delivery Network?Code9
KAG: Boosting LLMs in Professional Domains via Knowledge Augmented GenerationCode9
Language agents achieve superhuman synthesis of scientific knowledgeCode9
General OCR Theory: Towards OCR-2.0 via a Unified End-to-end ModelCode9
MaskGCT: Zero-Shot Text-to-Speech with Masked Generative Codec TransformerCode9
CogVLM2: Visual Language Models for Image and Video UnderstandingCode9
Sapiens: Foundation for Human Vision ModelsCode9
Transformer Explainer: Interactive Learning of Text-Generative ModelsCode9
SuperSimpleNet: Unifying Unsupervised and Supervised Learning for Fast and Reliable Surface Defect DetectionCode9
MindSearch: Mimicking Human Minds Elicits Deep AI SearcherCode9
NeedleBench: Can LLMs Do Retrieval and Reasoning in Information-Dense Context?Code9
MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse AttentionCode9
Diffusion Forcing: Next-token Prediction Meets Full-Sequence DiffusionCode9
Symbolic Learning Enables Self-Evolving AgentsCode9
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code IntelligenceCode9
Infinigen Indoors: Photorealistic Indoor Scenes using Procedural GenerationCode9
garak: A Framework for Security Probing Large Language ModelsCode9
BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-HaystackCode9
OpenVLA: An Open-Source Vision-Language-Action ModelCode9
Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image AnimationCode9
Show:102550
← PrevPage 7 of 20404Next →