NUMA-distributed weight banking for LLM inference on IBM POWER8. 147 t/s (8.8x stock). Part of the Proof of Physical AI stack.
transformer numa powerpc ppc64le power8 hebbian depin ai-inference llm llama-cpp llm-inference proof-of-physical-ai vec-perm
-
Updated
Aug 13, 2026 - Python