nvidia/foundationpose

NVIDIA releases FoundationPose, a unified transformer model for 6-DoF object pose estimation and tracking. The system accepts RGB, depth, and CAD inputs to estimate pose for novel objects without fine-tuning. The model card documents the architecture, license, and input formats, while the repository currently shows zero downloads, limiting immediate practical verification.

README

nvidia/foundationpose View on Hugging Face

Loading the README from Hugging Face…