Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Paper • 2608.30821 • Published 7 days ago • 109
DreamX-Creator: Democratizing Native Audio-Video Generation at 2K Resolution Paper • 2608.31106 • Published 7 days ago • 98
Round-Trip Consistency: Bidirectional Diffusion Models Can Predict Their Own Rollout Errors Paper • 2608.00675 • Published Aug 1 • 10
DreamActor-M2: Universal Character Image Animation via Spatiotemporal In-Context Learning Paper • 2601.21716 • Published Jan 29 • 13
Parallel Decoding Distillation for Fast Image and Video Generation Paper • 2607.26004 • Published Jul 28 • 17
JarvisHub: An Open Harness for Canvas-Native Multimodal Creative Agents Paper • 2607.23588 • Published Jul 26 • 126
DiffGI: Differentiable Geometry Images for High-Fidelity Thin-Shell 3D Generation Paper • 2607.13365 • Published Jul 15 • 22
Motion4Motion: Motion Transfer Across Subjects at Inference Paper • 2607.11644 • Published Jul 13 • 14
AlayaWorld: Long-Horizon and Playable Video World Generation Paper • 2607.06291 • Published Jul 7 • 92
SpatialAvatar-0: High-Quality 4D Head Avatar with Multi-Stage Reconstruction Paper • 2606.15659 • Published Jun 14 • 5
PermaVid: Consistent Video Generation Across Edits via Disentangled Context Memory Paper • 2606.16449 • Published Jun 15 • 7
World Tracing: Generative Pixel-Aligned Geometry Beyond the Visible Paper • 2606.13652 • Published Jun 11 • 17
Track2View: 4D-Consistent Camera-Controlled Video Generation via Paired 3D Point Tracks Paper • 2606.15534 • Published Jun 14 • 13
Memento: Reconstruct to Remember for Consistent Long Video Generation Paper • 2606.14667 • Published Jun 12 • 19
DreamX-World 1.0: A General-Purpose Interactive World Model Paper • 2606.16993 • Published Jun 15 • 115
Humanoid-GPT: Scaling Data and Structure for Zero-Shot Motion Tracking Paper • 2606.03985 • Published Jun 2 • 43