SOTAVerified

Gesture Generation

Generation of gestures, as a sequence of 3d poses

Papers

Showing 1–50 of 107 papers

TitleStatusHype
DeepGesture: A conversational gesture synthesis system based on emotions and semanticsCode0
Intentional Gesture: Deliver Your Intentions with Gestures for SpeechCode1
M3G: Multi-Granular Gesture Generator for Audio-Driven Full-Body Human Motion Synthesis—0
Inter-Diffusion Generation Model of Speakers and Listeners for Effective Communication—0
Co^3Gesture: Towards Coherent Concurrent Co-speech 3D Gesture Generation with Interactive Diffusion—0
EchoMask: Speech-Queried Attention-based Mask Modeling for Holistic Co-Speech Motion Generation—0
EasyGenNet: An Efficient Framework for Audio-Driven Gesture Video Generation Based on Diffusion Model—0
Audio-driven Gesture Generation via Deviation Feature in the Latent Space—0
SARGes: Semantically Aligned Reliable Gesture Generation via Intent Chain—0
DIDiffGes: Decoupled Semi-Implicit Diffusion Models for Real-time Gesture Generation from Speech—0
MAG: Multi-Modal Aligned Autoregressive Co-Speech Gesture Generation without Vector Quantization—0
Large Language Models for Virtual Human Gesture Selection—0
Streaming Generation of Co-Speech Gestures via Accelerated Rolling Diffusion—0
HOP: Heterogeneous Topology-based Multimodal Entanglement for Co-Speech Gesture Generation—0
VarGes: Improving Variation in Co-Speech 3D Gesture Generation via StyleCLIPSCode0
Contextual Gesture: Co-Speech Gesture Video Generation through Context-aware Gesture Representation—0
GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal ModelingCode2
EMO2: End-Effector Guided Audio-Driven Avatar Video Generation—0
Co-Speech Gesture Video Generation with Implicit Motion-Audio Entanglement—0
SemTalk: Holistic Co-speech Motion Generation with Frame-level Semantic Emphasis—0
The Language of Motion: Unifying Verbal and Non-verbal Language of 3D Human Motion—0
Retrieving Semantics from the Deep: an RAG Solution for Gesture SynthesisCode2
DiM-Gestor: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2—0
Conditional GAN for Enhancing Diffusion Models in Efficient and Authentic Global Gesture Generation from Audios—0
Large Body Language Models—0
Emphasizing Semantic Consistency of Salient Posture for Speech-Driven Gesture Generation—0
ExpGest: Expressive Speaker Generation Using Diffusion Model and Hybrid Audio-Text Guidance—0
Towards a GENEA Leaderboard -- an Extended, Living Benchmark for Evaluating and Advancing Conversational Motion Synthesis—0
LLM Gesticulator: Leveraging Large Language Models for Scalable and Controllable Co-Speech Gesture Synthesis—0
Enabling Synergistic Full-Body Control in Prompt-Based Co-Speech Motion GenerationCode0
MM-Conv: A Multi-modal Conversational Dataset for Virtual Humans—0
2D or not 2D: How Does the Dimensionality of Gesture Representation Affect 3D Co-Speech Gesture Generation?—0
Incorporating Spatial Awareness in Data-Driven Gesture Generation for Virtual Agents—0
MDT-A2G: Exploring Masked Diffusion Transformers for Co-Speech Gesture Generation—0
DiM-Gesture: Co-Speech Gesture Generation with Adaptive Layer Normalization Mamba-2 framework—0
MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal ControlsCode2
Investigating the impact of 2D gesture representation on co-speech gesture generation—0
AMUSE: Emotional Speech-driven 3D Body Animation via Disentangled Latent DiffusionCode2
CoCoGesture: Toward Coherent Co-speech 3D Gesture Generation in the Wild—0
LLAniMAtion: LLAMA Driven Gesture Animation—0
Bridge to Non-Barrier Communication: Gloss-Prompted Fine-grained Cued Speech Gesture Generation with Diffusion Model—0
ConvoFusion: Multi-Modal Conversational Diffusion for Co-Speech Gesture SynthesisCode1
Speech-driven Personalized Gesture Synthetics: Harnessing Automatic Fuzzy Feature Inference—0
MambaTalk: Efficient Holistic Gesture Synthesis with Selective State Space ModelsCode2
DiffSHEG: A Diffusion-Based Approach for Real-Time Speech-driven Holistic 3D Expression and Gesture Generation—0
Freetalker: Controllable Speech and Text-Driven Gesture Generation Based on Diffusion Models for Enhanced Speaker Naturalness—0
EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture ModelingCode3
Chain of Generation: Multi-Modal Gesture Synthesis via Cascaded Conditional Control—0
Emotional Speech-driven 3D Body Animation via Disentangled Latent DiffusionCode1
Weakly-Supervised Emotion Transition Learning for Diverse 3D Co-speech Gesture Generation—0
Show:102550
← PrevPage 1 of 3Next →

No leaderboard results yet.