Average Ratings 0 Ratings
Average Ratings 0 Ratings
Description
SAM 3D consists of a duo of sophisticated foundation models that can transform a typical RGB image into an impressive 3D representation of either objects or human figures. This system features SAM 3D Objects, which accurately reconstructs the complete 3D geometry, textures, and spatial arrangements of items found in real-world environments, effectively addressing challenges posed by clutter, occlusions, and varying lighting conditions. Additionally, SAM 3D Body generates dynamic human mesh models that capture intricate poses and shapes, utilizing the "Meta Momentum Human Rig" (MHR) format for enhanced detail. The design of this system allows it to operate effectively with images taken in natural settings without the need for further training or fine-tuning: users simply upload an image, select the desired object or individual, and receive a downloadable asset (such as .OBJ, .GLB, or MHR) that is instantly ready for integration into 3D software. Highlighting features like open-vocabulary reconstruction applicable to any object category, multi-view consistency, and occlusion reasoning, the models benefit from a substantial and diverse dataset containing over one million annotated images from the real world, which contributes significantly to their adaptability and reliability. Furthermore, the models are available as open-source, promoting wider accessibility and collaborative improvement within the development community.
Description
WorldClaw represents a comprehensive framework that facilitates the generation of expansive, freely navigable, and modifiable 3D open worlds, all derived from open-ended text prompts. Instead of constructing an entire world in one go, it employs a strategy that transitions from a broad overview to detailed regional features, ensuring spatial coherence while enriching local attributes. Initially, planning agents convert the text prompt into a structured scene specification that includes regions, terrain, assets, materials, visual aesthetics, and spatial arrangements. It establishes a globally consistent terrain base by utilizing semantic layouts, reusable assets, generative or procedural materials, and height fields that are sensitive to regional context. For areas that demand more intricate detail, WorldClaw generates terrain-conditioned compositions, reconstructs editable textured meshes, and appropriately places them within the scene. Subsequently, render-based agents enhance the terrain geometry, fine-tune object appearances, optimize arrangements, and ensure proper interactions with the surrounding environment. This multi-layered approach allows for both the creation of vast landscapes and the intricate detailing needed for immersive exploration.
API Access
Has API
API Access
Has API
Integrations
No details available.
Integrations
No details available.
Pricing Details
Free
Free Trial
Free Version
Pricing Details
Free
Free Trial
Free Version
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Deployment
Web-Based
On-Premises
iPhone App
iPad App
Android App
Windows
Mac
Linux
Chromebook
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Customer Support
Business Hours
Live Rep (24/7)
Online Support
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Types of Training
Training Docs
Webinars
Live Training (Online)
In Person
Vendor Details
Company Name
Meta
Founded
2004
Country
United States
Website
ai.meta.com/sam3d/
Vendor Details
Company Name
Tencent
Founded
1998
Country
China
Website
tencent-hunyuan.github.io/Hunyuan3D-WorldClaw/