Computer Vision

Detection, segmentation, OCR, and vision pipelines — production CV open source.

384 projects

See methodology for ranking rules; order uses public GitHub metrics within this scenario.

61–80 of 384

Rank Project Stars Forks
61 YOLOX

YOLOX is a high-performance anchor-free YOLO, exceeding yolov3~v5 with MegEngine, ONNX, TensorRT, ncnn, and OpenVINO supported. Documentation: https://yolox.readthedocs.io/

10.6K 2.5K
62 Final2x

a cross-platform image super-resolution tool

7.3K 525
63 doctr

docTR (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks powered by Deep Learning. Ongoing development and maintenance by t2k.

6.3K 673
64 caire

Content aware image resize library

10.5K 387
65 opencvsharp

OpenCV wrapper for .NET

6.1K 1.2K
66 jetson-inference

Hello AI World guide to deploying deep-learning inference networks and deep vision primitives with TensorRT and NVIDIA Jetson.

9.0K 3.1K
67 face-alignment

:fire: 2D and 3D Face alignment library build using pytorch

7.5K 1.4K
68 DeepLabCut

Official implementation of DeepLabCut: Markerless pose estimation of user-defined features with deep learning for all animals incl. humans

5.7K 1.8K
69 scrypted

Scrypted is a high performance video integration and automation platform

5.9K 391
70 PaddleX

All-in-One Development Tool based on PaddlePaddle

6.3K 1.2K
71 SlowFast

PySlowFast: video understanding codebase from FAIR for reproducing state-of-the-art video models.

7.4K 1.3K
72 sahi

Framework agnostic sliced/tiled inference + interactive ui + error analysis plots

5.5K 781
73 Pytorch-UNet

PyTorch implementation of the U-Net for image semantic segmentation with high quality images

11.6K 2.8K
74 clip-as-service

🏄 Scalable embedding, reasoning, ranking for images and sentences with CLIP

12.8K 2.1K
75 Edit-Banana

Edit Banana: A framework for converting statistical formats into editable.

5.5K 358
76 facenet

Face recognition using Tensorflow

14.3K 4.8K
77 remove-ai-watermarks

Remove visible and invisible AI watermarks and provenance metadata from images and video. Python library and CLI for SynthID, C2PA, EXIF, IPTC, XMP, and common generative-AI marks.

5.3K 499
78 gemini-watermark-remover

A high-performance, 100% client-side tool for removing Gemini AI image & video watermarks. Built with pure JavaScript using mathematically precise Reverse Alpha Blending. / 基于 JavaScript 的纯浏览器端 Gemini AI 图像和视频无损去水印工具,使用数学精确的反向 Alpha 混合算法

5.4K 861
79 sports

computer vision and sports

5.3K 656
80 MaaFramework

An automation black-box testing framework based on image recognition

4.7K 543