VGGT-MPS
Provides 3D vision and reconstruction capabilities through VGGT transformer model for multi-view reconstruction from image sequences, offering camera pose estimation, depth prediction, point cloud generation, and point tracking with Apple Silicon MPS acceleration for efficient inference on Mac hardw
About
This VGGT-MPS server provides 3D vision and reconstruction capabilities optimized for Apple Silicon through MPS acceleration, implementing the VGGT (Vision-based Geometry and Tracking Transformer) model for multi-view 3D reconstruction from image sequences. Built with PyTorch and FastMCP, it offers tools for camera pose estimation, depth prediction, 3D point cloud generation, and point tracking across video frames, featuring specialized MPS optimizations for efficient inference on Apple's Metal Performance Shaders framework. The implementation includes Gradio web interfaces, COLMAP integration for structure-from-motion workflows, and sparse attention mechanisms for scaling to large image collections, making it valuable for researchers and developers working on 3D computer vision applications who need GPU-accelerated reconstruction on Mac hardware without CUDA dependencies.
Is this your project?
Claim this listing to manage your page, access analytics, and unlock upgrades. Verification takes 60 seconds.
Share This Project
Embed Badge
Add this badge to your README:
[](https://hifriendbot.com/ai-list/vggt-mps/)
