Popular repositories Loading
-
q36
q36 PublicRun Qwen 3.6 35B on RTX 5090 hardware using a zero-dependency C/CUDA inference engine designed for peak throughput and low latency.
Cuda
-
rochelleveritable865.github.io
rochelleveritable865.github.io PublicAccelerate Qwen3.6 35B inference on RTX 5090 hardware using this hyper-optimized, zero-dependency C/CUDA engine for superior prefill and decode speeds.
HTML
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.