Built on CUDA 11.5 and run several NVIDIA architectures on 8.6 (Ampere) GPU (runtime compile of virtuals takes minutes, empty images on virtual Volta and below) #2082
ProbablyMartian
started this conversation in
Benchmark
Replies: 0 comments
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
I have tried to build and run all architectures as separate
sd-clifiles - from ones that got included (on my system) when following default CUDA build instructions - from 50 (Maxwell) to 86 (Ampere).Disclaimer: tests done on one GPU (Ampere) + CPU (no AVX2), on one model (Z-Image-Turbo Q4/Q8) and one CUDA version (11.5), OS (Linux Mint 21 aka LM21). If your observations show different picture, please let me know.
Observations:
I suggest
sd.cppadd some line indebugmode (or even higher) when starting to compile virtual CUDA.So hopes of forward-compatibility go out the window, correct? Well, at least for CUDA, how to check Vulkan? I guess we have no
sd.cppbuilds from Maxwell days, do we?I have made other minor observations, if you are interested in more details, let me know.
Details on testing:
sd.cppgit downloaded some time September 2026.for f in CUDA-folder/sd-cli-*; do for (( i=1;i<=3;i++ )) do~/.nv/ComputeCachebetween loops runs (I have not usedCUDA_FORCE_PTX_JIT=1 sdi-clicause I wanted to compare with compile time first to next compiled time runs).P.S. related: I did build and run on CUDA 12 (LM 22) too a bit on my Ampere RTX. On LM21 - Vulkan's sd.cpp release does not work, but works on LM22. Speed of CUDA's is ~1.2x of Vulkan's, but I liked results (images) of Vulkan's produced better as same settings.
All reactions