Atlas by World LabsWorld Labs
|
ModelScopeAlibaba Cloud
|
|||||
Related Products
|
||||||
About
Atlas is a next-generation omni world model for spatial intelligence that natively operates on text, images, video, and 3D. Built as a multimodal autoregressive diffusion transformer, it combines inputs into a shared spatial context and generates what comes next while staying consistent in 3D with everything it has seen and imagining what lies beyond it. Atlas supports world generation, reconstruction, and simulation across a broad range of tasks. It can generate images and videos from one or more reference images with pixel-perfect camera control, producing long, coherent videos with manually designed camera paths. For spatial reconstruction, Atlas can recreate real-world scenes from sparse input images, generate novel views, and produce explicit 3D outputs such as point clouds and 3D Gaussian splats. More input views provide additional context, allowing the model to reduce imagination and create increasingly faithful reconstructions. Atlas also models how worlds evolve over time.
|
About
This model is based on a multi-stage text-to-video generation diffusion model, which inputs a description text and returns a video that matches the text description. Only English input is supported.
This model is based on a multi-stage text-to-video generation diffusion model, which inputs a description text and returns a video that matches the text description. Only English input is supported.
The text-to-video generation diffusion model consists of three sub-networks: text feature extraction, text feature-to-video latent space diffusion model, and video latent space to video visual space. The overall model parameters are about 1.7 billion. Support English input. The diffusion model adopts the Unet3D structure, and realizes the function of video generation through the iterative denoising process from the pure Gaussian noise video.
|
|||||
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
Platforms Supported
Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook
|
|||||
Audience
Creators, VFX artists, robotics teams, game developers, designers, and AI researchers in need of a tool to generate, reconstruct, and simulate spatially consistent 3D worlds from multimodal inputs
|
Audience
Users interested in an open source text-to-video AI video generation model
|
|||||
Support
Phone Support
24/7 Live Support
Online
|
Support
Phone Support
24/7 Live Support
Online
|
|||||
API
Offers API
|
API
Offers API
|
|||||
Screenshots and Videos |
Screenshots and Videos |
|||||
Pricing
No information available.
Free Version
Free Trial
|
Pricing
Free
Free Version
Free Trial
|
|||||
Reviews/
|
Reviews/
|
|||||
Training
Documentation
Webinars
Live Online
In Person
|
Training
Documentation
Webinars
Live Online
In Person
|
|||||
Company InformationWorld Labs
Founded: 2024
United States
www.worldlabs.ai/blog/atlas
|
Company InformationAlibaba Cloud
China
modelscope.cn/
|
|||||
Alternatives |
Alternatives |
|||||
|
|
||||||
|
|
||||||
|
|
||||||
|
|
|
|||||
Categories |
Categories |
|||||
Integrations
01.AI
CodeQwen
Qwen
Qwen-7B
Qwen-Image
Qwen2-VL
Qwen2.5
Qwen2.5-1M
Qwen2.5-Coder
Qwen2.5-Max
|
Integrations
01.AI
CodeQwen
Qwen
Qwen-7B
Qwen-Image
Qwen2-VL
Qwen2.5
Qwen2.5-1M
Qwen2.5-Coder
Qwen2.5-Max
|
|||||
|
|
|