llama-bench Strix Halo (gfx1151) ROCm 7.2.3 + Qwen3.5 122B A10B MTP UD-Q6_K_XL
#23659
muslimpribadi
started this conversation in
Show and tell
Replies: 0 comments
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Uh oh!
There was an error while loading. Please reload this page.
Uh oh!
There was an error while loading. Please reload this page.
Run
llama-benchusing unsloth/Qwen3.5-122B-A10B-MTP-GGUFUD-Q6_K_XL.Header llama-bench output:
Footer llama-bench output:
Text generation Qwen3.5 122B A10B with gemma 4 31B (dense)
🟢 Benchmark result
Prompt processing with different batch sizes
🟢 Benchmark result
Different numbers of threads
🟢 Benchmark result
Different numbers of ubatch and threads
🟢 Benchmark result
Environment
Container:
podmanimagedocker pull kyuz0/amd-strix-halo-toolboxes:rocm-7.2.3ROCM 7.2.3All reactions