⚡ Self-registering node agent for distributed AI clusters, dynamically serving local LLMs via vLLM across heterogeneous GPUs.
distributed-systems cuda hardware-detection gpu-cluster node-agent tailscale zero-touch-provisioning vllm local-llm llm-inference gpu-worker
-
Updated
Sep 20, 2026 - Python