Skip to content

MultiCyclone v3.1 - multi-architecture (Turing and up)

Choose a tag to compare

@kemo159 kemo159 released this 27 Aug 19:16
· 3 commits to main since this release

Same code as v3.1
— only the compiled GPU architectures differ. Grab these if you are not on an
RTX 50-series card.

architectures GPUs size
this release sm_75, 80, 86, 89, 90, 120 Turing → Blackwell 41–54 MB
v3.1 sm_120 only RTX 50-series only 9–15 MB

That covers RTX 20 / GTX 16 (Turing), A100, RTX 30 (Ampere), RTX 40 (Ada), H100
(Hopper) and RTX 50 (Blackwell).

asset built with needs
CUDACyclone-v3.1-windows-x64-multiarch.exe CUDA 13.1 + MSVC 14.44 NVIDIA driver for CUDA 13.x
CUDACyclone-v3.1-linux-x64-multiarch CUDA 12.8 + g++, Ubuntu 24.04 NVIDIA driver for CUDA 12.8+, chmod +x

The Linux build is deliberately on CUDA 12.8 rather than 13.x: it asks for an
older minimum driver, so it runs on machines that have not updated.

There is no speed penalty. A fatbinary carries native SASS per architecture,
so the card runs its own code either way. The only cost is file size.

What is new in v3.1

--autosavetimer SECONDS rewrites the checkpoint every N seconds while the
search keeps running, so a crash, a power cut or kill -9 costs at most one
interval instead of the whole run. Ctrl+C and --seconds already saved on the
way out; this covers the stops that never get the chance.

CUDACyclone --range AAAA:BBBB --address 1Abc... --autosavetimer 300
[autosave] checkpoint written to cyclone_checkpoint.txt at 12.04% (48318382080 keys checked)

Recovery is the ordinary --resume path. The save does not pause the search, it
cannot skip keys, and a crash mid-save cannot truncate the previous good
checkpoint — see v3.1
for the details.

Anything older than Turing

Not covered. CUDA 13.x removed Maxwell, Pascal and Volta entirely, and while
CUDA 12.8 can still target them, none of it is tested here. Build from source if
you need one.