Inference & Local Deploy

High-throughput serving and local runtimes — Ollama, vLLM, Triton, and more.

183 projects

See methodology for ranking rules; order uses public GitHub metrics within this scenario.

181–183 of 183

Rank Project Stars Forks
181 tensorRT_Pro

C++ library based on tensorrt integration

2.9K 573
182 keras-js

Run Keras models in the browser, with GPU support using WebGL

5.0K 491
183 keras_to_tensorflow

General code to convert a trained keras model into an inference tensorflow model

1.6K 524
< Previous
/ 10
Next >