#88 · Primary category: Inference & Local Deploy
inference
Turn any computer or edge device into a command center for your computer vision projects.
Project last updated:08/28/26
GitHub Stars
2.4K
Forks
311
Contributors
122
License
Other
Why we included this project
Roboflow's inference server lets you run computer vision models on hardware you control, handling the serving details so you can focus on the model itself. Behind one REST API and Python SDK sit YOLO-family detectors, segmentation, classification, and foundation models like Florence-2, CLIP, and SAM2, so swapping models doesn't touch your application. The Workflows feature chains detection, tracking, counting, OCR, and your own logic into pipelines that respond to live camera feeds, which covers real tasks like reading license plates or alerting when occupancy changes. Docker images and Jetson support make GPU-accelerated deployment on edge devices straightforward.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
ollama
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
llama.cpp
LLM inference in C/C++
vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
gpt4all
GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.
LocalAI
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.