#133 · Primary category: Computer Vision
vipe
ViPE: Video Pose Engine for Geometric 3D Perception
Project last updated:08/17/26
GitHub Stars
2.1K
Forks
172
Contributors
11
License
Other
Why we included this project
ViPE extracts camera intrinsics, motion, and dense near-metric depth from raw video, including wide-angle and 360-degree panorama footage. That covers the geometry most 3D reconstruction and SLAM pipelines need, without forcing you to stitch together separate tools. The Python package installs from PyPI and comes with documentation and dataset support, so it fits into existing annotation workflows. Recent releases added CUDA-fused kernels and caching that speed things up roughly 2.7x with no loss of accuracy. Teams that need ground-truth poses and depth for training or evaluation can treat this as a dependable source.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
opencv
Open Source Computer Vision Library
RuView
π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
MinerU
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
tesseract
Tesseract Open Source OCR Engine (main repository)