Inference & Local Deploy
High-throughput serving and local runtimes — Ollama, vLLM, Triton, and more.
183 projects
See methodology for ranking rules; order uses public GitHub metrics within this scenario.
| Rank | Project | Stars | Forks | Updated | License |
|---|---|---|---|---|---|
| 181 |
tensorRT_Pro
C++ library based on tensorrt integration |
2.9K | 573 | 05/24/23 | MIT |
| 182 |
keras-js
Run Keras models in the browser, with GPU support using WebGL |
5.0K | 491 | 06/15/22 | MIT |
| 183 |
keras_to_tensorflow
General code to convert a trained keras model into an inference tensorflow model |
1.6K | 524 | 11/23/20 | MIT |