#118 · Primary category: Computer Vision
DeepCamera
Open-source AI camera skills platform with local VLM analysis and agentic security agent for home surveillance via messaging apps.
Project last updated:06/18/26
GitHub Stars
3.0K
Forks
477
Contributors
21
License
MIT
Why we included this project
DeepCamera is for people who run home or small-business cameras and want the feeds to do more than just record. It packages the usual surveillance work, object detection, face recognition, license-plate reading, depth-based anonymization, and SAM2 segmentation, into skills that all run on your own hardware, so footage never leaves the premises. Scene understanding and fall detection lean on modern vision-language models, and alerts can be pushed to Telegram, Discord, or Slack. The same skill also runs on NVIDIA, Apple Silicon, Intel, or a Coral USB TPU with automatic model conversion, which makes the hardware choice less of a commitment. If you want a pluggable surveillance setup with local inference and the ability to add new capabilities without rebuilding the core, this is a practical starting point.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
opencv
Open Source Computer Vision Library
RuView
π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
MinerU
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
tesseract
Tesseract Open Source OCR Engine (main repository)