Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
Atlas represents an advanced omni world model designed for spatial intelligence, seamlessly engaging with text, images, video, and 3D data. As a sophisticated multimodal autoregressive diffusion transformer, it integrates various inputs into a common spatial framework while predicting subsequent elements and ensuring coherence in 3D with its observations and creative visions. This powerful model facilitates world generation, reconstruction, and simulation across diverse applications. It possesses the capability to produce images and videos based on one or multiple reference images, offering precise camera control to yield lengthy and coherent videos following custom-designed camera trajectories. In terms of spatial reconstruction, Atlas excels at reimagining real-world environments from limited input images, allowing it to generate unique perspectives and deliver explicit 3D outputs like point clouds and 3D Gaussian splats. With the addition of more input views, the model can further enhance context, leading to increasingly accurate reconstructions with less reliance on imagination. Furthermore, Atlas also analyzes and predicts the evolution of worlds over time, adding a dynamic layer to its capabilities.
Description
SAM 3D consists of a duo of sophisticated foundation models that can transform a typical RGB image into an impressive 3D representation of either objects or human figures. This system features SAM 3D Objects, which accurately reconstructs the complete 3D geometry, textures, and spatial arrangements of items found in real-world environments, effectively addressing challenges posed by clutter, occlusions, and varying lighting conditions. Additionally, SAM 3D Body generates dynamic human mesh models that capture intricate poses and shapes, utilizing the "Meta Momentum Human Rig" (MHR) format for enhanced detail. The design of this system allows it to operate effectively with images taken in natural settings without the need for further training or fine-tuning: users simply upload an image, select the desired object or individual, and receive a downloadable asset (such as .OBJ, .GLB, or MHR) that is instantly ready for integration into 3D software. Highlighting features like open-vocabulary reconstruction applicable to any object category, multi-view consistency, and occlusion reasoning, the models benefit from a substantial and diverse dataset containing over one million annotated images from the real world, which contributes significantly to their adaptability and reliability. Furthermore, the models are available as open-source, promoting wider accessibility and collaborative improvement within the development community.
API Access
Has API
Yes
API Access
Has API
No
Integrations
No details available.
Integrations
No details available.
Pricing Details
No price information available.
Free Trial
No
Free Version
No
Pricing Details
Free
Free Trial
Yes
Free Version
Yes
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Deployment
Web-Based
Yes
On-Premises
No
iPhone App
No
iPad App
No
Android App
No
Windows
No
Mac
No
Linux
No
Chromebook
No
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Customer Support
Business Hours
No
Live Rep (24/7)
No
Online Support
Yes
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
No
In Person
No
Types of Training
Training Docs
Yes
Webinars
No
Live Training (Online)
Yes
In Person
No
Vendor Details
Company Name
World Labs
Founded
2024
Country
United States
Website
www.worldlabs.ai/blog/atlas
Vendor Details
Company Name
Meta
Founded
2004
Country
United States
Website
ai.meta.com/sam3d/