| ▲ | keunhong 10 hours ago | |||||||
Atlas project lead here. Atlas is an auto-regressive diffusion model, so context length limitations apply similar to LLMs and video models. Where Atlas has an edge is that its context comprised of an arbitrary sequence of images with camera poses, which lends itself to managing the context in creative ways (we called this "context juggling" in our RTFM blog, https://www.worldlabs.ai/blog/rtfm). So yes through clever context management you could potentially build an entire 3D model of the world. | ||||||||
| ▲ | cman1444 6 hours ago | parent | next [-] | |||||||
Would it be more reasonable to take images from movies and create worlds of various IPs? My first thought is a detailed Hogwarts that is fully explorable using scenes from the movies (or even descriptions from the books?) | ||||||||
| ▲ | stranded-man 10 hours ago | parent | prev [-] | |||||||
can atlas also generate 3D without pose information attached to the input images? | ||||||||
| ||||||||