SOTAVerified

The Open Verification Layer for ML Research

Community benchmark tracking and reproducibility verification. Built for researchers and autonomous research agents.

510,095 papers251,776 code links4,818 tasks

Papers

Showing 451500 of 177341 papers

TitleStatusHype
PIXART-δ: Fast and Controllable Image Generation with Latent Consistency ModelsCode7
Large Language Model Agent: A Survey on Methodology, Applications and ChallengesCode7
Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion TransformersCode7
SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the WildCode7
Logo-LLM: Local and Global Modeling with Large Language Models for Time Series ForecastingCode7
DragAnything: Motion Control for Anything using Entity RepresentationCode7
EvoRL: A GPU-accelerated Framework for Evolutionary Reinforcement LearningCode7
Efficient MedSAMs: Segment Anything in Medical Images on LaptopCode7
Aligning Anime Video Generation with Human FeedbackCode7
Chronos: Learning the Language of Time SeriesCode7
Adding Conditional Control to Text-to-Image Diffusion ModelsCode7
OASIS: Open Agent Social Interaction Simulations with One Million AgentsCode7
Muon is Scalable for LLM TrainingCode7
An Empirical Study on Reinforcement Learning for Reasoning-Search Interleaved LLM AgentsCode7
Step-Audio-AQAA: a Fully End-to-End Expressive Large Audio Language ModelCode7
EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark ConditionsCode7
CodexGraph: Bridging Large Language Models and Code Repositories via Code Graph DatabasesCode7
Adaptive In-conversation Team Building for Language Model AgentsCode7
Visual ChatGPT: Talking, Drawing and Editing with Visual Foundation ModelsCode7
MiniMax-01: Scaling Foundation Models with Lightning AttentionCode7
Drag Your GAN: Interactive Point-based Manipulation on the Generative Image ManifoldCode7
MeshAnything: Artist-Created Mesh Generation with Autoregressive TransformersCode7
BricksRL: A Platform for Democratizing Robotics and Reinforcement Learning Research and Education with LEGOCode7
EAGLE-2: Faster Inference of Language Models with Dynamic Draft TreesCode7
DeepSpeed-FastGen: High-throughput Text Generation for LLMs via MII and DeepSpeed-InferenceCode7
MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio SynthesisCode7
AniSora: Exploring the Frontiers of Animation Video Generation in the Sora EraCode7
Better than classical? The subtle art of benchmarking quantum machine learning modelsCode7
Ichigo: Mixed-Modal Early-Fusion Realtime Voice AssistantCode7
GenAD: Generalized Predictive Model for Autonomous DrivingCode7
FAST-LIVO2: Fast, Direct LiDAR-Inertial-Visual OdometryCode7
aiXcoder-7B: A Lightweight and Effective Large Language Model for Code ProcessingCode7
AgentOrchestra: A Hierarchical Multi-Agent Framework for General-Purpose Task SolvingCode7
MAGI-1: Autoregressive Video Generation at ScaleCode7
Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese UnderstandingCode7
ComfyUI-Copilot: An Intelligent Assistant for Automated Workflow DevelopmentCode7
Kimi-Audio Technical ReportCode7
Bilateral Reference for High-Resolution Dichotomous Image SegmentationCode7
EvoGP: A GPU-accelerated Framework for Tree-based Genetic ProgrammingCode7
AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied SystemsCode7
StarCoder 2 and The Stack v2: The Next GenerationCode7
Mini-Omni: Language Models Can Hear, Talk While Thinking in StreamingCode7
Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning SystemsCode7
Neural Codec Language Models are Zero-Shot Text to Speech SynthesizersCode7
DocETL: Agentic Query Rewriting and Evaluation for Complex Document ProcessingCode7
Intent-based Prompt Calibration: Enhancing prompt optimization with synthetic boundary casesCode7
Improving Sample Quality of Diffusion Models Using Self-Attention GuidanceCode7
EasyAnimate: A High-Performance Long Video Generation Method based on Transformer ArchitectureCode7
HunyuanVideo-Avatar: High-Fidelity Audio-Driven Human Animation for Multiple CharactersCode7
MagicQuill: An Intelligent Interactive Image Editing SystemCode7
Show:102550
← PrevPage 10 of 3547Next →