#162 · Primary category: Inference & Local Deploy
Talos
GPU worker client for the Talos network. Pairs with your Talos account, serves open-model inference jobs over a WebSocket, and reports uptime for payouts.
Project last updated:07/08/26
GitHub Stars
712
Forks
18
Contributors
1
License
MIT
Why we included this project
This lightweight Python client lets you put an idle GPU to work. You pair it to a Talos account with a short code, and it listens over a WebSocket for open-model inference jobs, running them through your local Ollama install. An allocation slider sets how much of the machine you offer, and the dashboard tracks uptime and per-job earnings. If you already run Ollama and want to earn from spare capacity without standing up your own serving stack, this is a low-effort way to join a distributed inference network.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
ollama
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
llama.cpp
LLM inference in C/C++
vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
gpt4all
GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.
LocalAI
LocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.