-
Notifications
You must be signed in to change notification settings - Fork 5
Home
youngharold edited this page Feb 17, 2026
·
12 revisions
Mixed-vendor GPU inference cluster manager using llama.cpp RPC.
- Architecture — How the cluster works, data flow, and design decisions
- Hardware Setup — Building llama.cpp for CUDA/ROCm workers and coordinator
- Configuration — cluster.yaml reference and examples
- CLI Reference — All commands and options
- Troubleshooting — Common issues and fixes
- Network Optimization — Bandwidth tuning and layer placement
- OpenClaw Integration — Registering Hydra as an OpenClaw provider