nvidia/foundationpose
NVIDIA releases FoundationPose, a unified transformer model for 6-DoF object pose estimation and tracking. The system accepts RGB, depth, and CAD inputs to estimate pose for novel objects without fine-tuning. The model card documents the architecture, license, and input formats, while the repository currently shows zero downloads, limiting immediate practical verification.
README
nvidia/foundationpose View on Hugging Face
Loading the README from Hugging Face…