171 upvotes · 11 AUG 2026 · Zihao Liu, Xiaolong Shen, Zhenglin Zhou et al.
This paper develops a method to generate 4D worlds from text or images by leveraging the latent space of a video prior. It allows for more flexible and reusable 4D prediction, enabling practitioners to create more realistic and stable 3D scenes.
58 upvotes · 5 AUG 2026 · Chunchao Guo, Jinpeng Li, Yang Li et al.
This paper develops a system for generating large-scale, open-world 3D environments from text prompts, with a focus on coherence and editability. Practitioners might care about the potential applications of this technology in fields like game development, architecture, and visual effects.