#110 · Primary category: Computer Vision
paperless-gpt
Use LLMs and LLM Vision (OCR) to handle paperless-ngx - Document Digitalization powered by AI
Project last updated:08/28/26
GitHub Stars
2.7K
Forks
204
Contributors
32
License
MIT
Why we included this project
People who run paperless-ngx with a large scanned archive will find this add-on useful because it treats OCR as a real bottleneck. paperless-gpt connects to your existing instance and uses LLM vision models like OpenAI and Ollama, plus cloud services such as Azure and Google Document AI, to turn low-quality scans into context-aware text, then writes that back as a searchable PDF text layer. It also suggests document titles, tags, correspondents, and custom field values, and gives you a web UI to review each suggestion before it's applied. That hands-on control, combined with better-than-traditional OCR, makes it a step beyond auto-tagging-only tools, and the Docker deployment keeps self-hosting straightforward.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
opencv
Open Source Computer Vision Library
RuView
π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
PaddleOCR
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
MinerU
Transforms complex documents like PDFs and Office docs into LLM-ready markdown/JSON for your Agentic workflows.
tesseract
Tesseract Open Source OCR Engine (main repository)