Motion-Omni: End-to-End Joint Speech and Full-Body Motion for Spoken Dialogue Paper • 2609.04250 • Published 18 days ago • 42
Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Paper • 2608.30821 • Published 15 days ago • 121
DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution Paper • 2608.31106 • Published 15 days ago • 101
Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors Paper • 2608.00675 • Published Aug 1 • 10
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning Paper • 2601.21716 • Published Jan 29 • 13
Parallel Decoding Distillation for Fast Image and Video Generation Paper • 2607.26004 • Published Jul 28 • 18
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published Jul 26 • 127
DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation Paper • 2607.13365 • Published Jul 15 • 23
Motion4Motion: Motion Transfer Across Subjects at Inference Paper • 2607.11644 • Published Jul 13 • 15
AlayaWorld: Long-Horizon and Playable Video World Generation Paper • 2607.06291 • Published Jul 7 • 93
SpatialAvatar-0: High-Quality 4D Head Avatar with Multi-Stage Reconstruction Paper • 2606.15659 • Published Jun 14 • 5
PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory Paper • 2606.16449 • Published Jun 15 • 7
World Tracing: Generative Pixel-Aligned Geometry Beyond the Visible Paper • 2606.13652 • Published Jun 11 • 17
Track2View: 4D-Consistent Camera-Controlled Video Generation via Paired 3D Point Tracks Paper • 2606.15534 • Published Jun 14 • 13
Memento: Reconstruct to Remember for Consistent Long Video Generation Paper • 2606.14667 • Published Jun 12 • 19
DreamX-World 1.0: A General-Purpose Interactive World Model Paper • 2606.16993 • Published Jun 15 • 116