Computer Vision
Detection, segmentation, OCR, and vision pipelines — production CV open source.
452 projects
See methodology for ranking rules; order uses public GitHub metrics within this scenario.
| Rank | Project | Stars | Forks | Updated | License |
|---|---|---|---|---|---|
| 381 |
RingNet
Learning to Regress 3D Face Shape and Expression from an Image without 3D Supervision |
884 | 172 | 03/24/23 | MIT |
| 382 |
CRNN_Chinese_Characters_Rec
(CRNN) Chinese Characters Recognition. |
1.9K | 533 | 11/13/22 | Other |
| 383 |
interactive-deep-colorization
Deep learning software for colorizing black and white images with a few clicks. |
2.7K | 446 | 07/29/22 | MIT |
| 384 |
lip-reading-deeplearning
:unlock: Lip Reading - Cross Audio-Visual Recognition using 3D Architectures |
1.9K | 331 | 11/07/22 | Apache-2.0 |
| 385 |
image-to-latex
Convert images of LaTex math equations into LaTex code. |
2.2K | 311 | 10/04/22 | MIT |
| 386 |
pytorch-center-loss
Pytorch implementation of Center Loss |
993 | 216 | 02/19/23 | MIT |
| 387 |
Medical-Transformer
Official Pytorch Code for "Medical Transformer: Gated Axial-Attention for Medical Image Segmentation" - MICCAI 2021 |
861 | 175 | 02/23/23 | MIT |
| 388 |
flame-fitting
Example code for the FLAME 3D head model. The code demonstrates how to sample 3D heads from the model, fit the model to 3D keypoints and 3D scans. |
820 | 120 | 02/16/23 | Other |
| 389 |
QuickDraw
Implementation of Quickdraw - an online game developed by Google |
1.1K | 195 | 01/11/23 | MIT |
| 390 |
tencent-ml-images
Largest multi-label image database; ResNet-101 model; 80.73% top-1 acc on ImageNet |
3.1K | 503 | 04/20/22 | Other |
| 391 |
grad-cam
[ICCV 2017] Torch code for Grad-CAM |
1.7K | 237 | 09/17/22 | Other |
| 392 |
PWC-Net
PWC-Net: CNNs for Optical Flow Using Pyramid, Warping, and Cost Volume, CVPR 2018 (Oral) |
1.7K | 365 | 08/22/22 | Other |
| 393 |
flamingo-pytorch
Implementation of 🦩 Flamingo, state-of-the-art few-shot visual question answering attention net out of Deepmind, in Pytorch |
1.3K | 66 | 10/18/22 | MIT |
| 394 |
redner
Differentiable rendering without approximation. |
1.4K | 145 | 08/19/22 | MIT |
| 395 |
Yolov5-Deepsort
Latest YOLOv5+DeepSORT for object detection and tracking, displays categories, supports v5.0 and custom dataset training. |
1.2K | 171 | 10/06/22 | GPL-3.0 |
| 396 |
PaddleViT
:robot: PaddleViT: State-of-the-art Visual Transformer and MLP Models for PaddlePaddle 2.0+ |
1.2K | 328 | 09/07/22 | Apache-2.0 |
| 397 |
natural-language-image-search
Search photos on Unsplash using natural language |
1.0K | 103 | 10/13/22 | MIT |
| 398 |
keras-vis
Neural network visualization toolkit for keras |
3.0K | 635 | 02/07/22 | MIT |
| 399 |
gangealing
Official PyTorch Implementation of "GAN-Supervised Dense Visual Alignment" (CVPR 2022 Oral, Best Paper Finalist) |
1.0K | 121 | 10/12/22 | BSD-2-Clause |
| 400 |
captcha_break
Captcha recognition |
2.8K | 667 | 02/25/22 | MIT |