Ben Mildenhall@BenMildenhallour new model Atlas also happens to be a capable text-to-image generator, providing it with a strong foundation of “world knowledge”... this generalizes across both single and multi view domains, meaning you can step into even highly stylized scenes with precise 3D controlOpens with an observation
Ben Mildenhall@BenMildenhallAtlas is an autoregressive diffusion model built from the ground up for the task of "next frame prediction". It is simultaneously a world class method for camera-controlled video generation, novel view synthesis, and sparse 3D reconstruction.Opens with an observation