World Labs has launched Atlas, a new world model that unifies pixel generation and 3D reconstruction through a novel 'new view prediction' primitive, enabling high-fidelity spatial intelligence with significantly reduced data requirements. The model achieves 50-100x reductions in capture effort by using sparse inputs (as few as 3 cameras) to generate dense 3D environments, impacting creative workflows and robotics simulation. Future development focuses on scaling compute, enhancing dynamics, and improving editability to bridge the gap between simulation and real-world robotic deployment.
“We know LLMs are built on next token prediction. We've seen video models as being built on next frame prediction. Atlas is really new view prediction.”
“There was a famous shot in the first Matrix movie where Neo is like falling down... They had hundreds of cameras viewing that angle on a green screen. On Atlas, we can do this with just three cameras. No studio capture, no green screen, no expensive calibration.”
“This is a elegant model that combines or unifies the problem of reconstruction and generation by anchoring on viewpoints... that's just incredibly powerful.”
“Intelligence is not sitting there stack and just seeing something or interpreting something when it comes to space and physical space... It's really this closing the loop between seeing and experiencing and interaction.”
“New view prediction... is also AI complete... because I could take something like you could have the movie and you do all of the frames of the movie and then like the killer walks out and then you predict exactly who walks out.”
56m
1h 4m
53m
50m
58m
47m
50m
45m
1h 3m
1h 8m
1h 3m
56m
1h 6m⏭ 2
1h 44m⏭ 3
1h 2m⏭ 2
2h 40m⏭ 4
2h 11m⏭ 3
1h 1m⏭ 1
2h 57m
2h 43m
1h 46m⏭ 3
45m
2h 0m
1h 43m⏭ 1
11h 23m⏭ 1
33m
1h 0m⏭ 1
46m
3h 54m⏭ 7
48m⏭ 2
3h 59m⏭ 5
1h 13m
1h 16m
54m⏭ 1
58m⏭ 5
24m