Build the GPU inference stack from source, repeatably: for Blackwell sm_120 on CUDA 13.x and Python 3.14. Machine-readable build state, generated patches, and an abductive-triage skill for build failures.
gpu cuda inference pytorch nvidia blackwell build-from-source llama-cpp vllm agent-skills sm120 python-314 cuda-13 abductive-triage
-
Updated
Aug 4, 2026 - Shell