Computer Vision

Detection, segmentation, OCR, and vision pipelines — production CV open source.

384 projects

See methodology for ranking rules; order uses public GitHub metrics within this scenario.

1–20 of 384

Rank Project Stars Forks
1 opencv

Open Source Computer Vision Library

90.6K 57.0K
2 RuView

π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.

92.0K 12.2K
3 PaddleOCR

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

88.5K 11.3K
4 MinerU

Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.

78.7K 6.6K
5 tesseract

Tesseract Open Source OCR Engine (main repository)

76.2K 10.8K
6 ultralytics

Ultralytics YOLO26, YOLO11, YOLOv8 — object detection, instance segmentation, semantic segmentation, image classification, pose estimation, object tracking

61.1K 11.7K
7 yolov5

Ultralytics YOLOv5 in PyTorch for object detection, instance segmentation, classification, training, and export.

57.9K 17.5K
8 faceswap

Deepfakes Software For All

57.5K 13.5K
9 face_recognition

The world's simplest facial recognition api for Python and the command line

56.7K 13.7K
10 supervision

We write your reusable computer vision tools. 💜

49.8K 4.7K
11 photoprism

AI-Powered Photos App 🌈💎✨

40.1K 2.3K
12 tesseract.js

Pure Javascript OCR for more than 100 Languages 📖🎉🖥

38.7K 2.4K
13 frigate

NVR with realtime local object detection for IP cameras

35.5K 3.5K
14 mediapipe

Cross-platform, customizable ML solutions for live and streaming media.

36.8K 6.1K
15 GFPGAN

GFPGAN aims at developing Practical Algorithms for Real-world Face Restoration.

37.7K 6.3K
16 Real-ESRGAN

Real-ESRGAN aims at developing Practical Algorithms for General Image/Video Restoration.

36.6K 4.5K
17 facefusion

Industry leading face manipulation platform

29.7K 4.8K
18 openpose

OpenPose: Real-time multi-person keypoint detection library for body, face, hands, and foot estimation

34.4K 8.0K
19 EasyOCR

Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.

30.0K 3.6K
20 vit-pytorch

Implementation of Vision Transformer, a simple way to achieve SOTA in vision classification with only a single transformer encoder, in Pytorch

25.5K 3.5K
< Previous
/ 20
Next >