Observed
Against main da2291339 (first executed on PR #2370's merge ref; main's own run had these jobs cancelled before executing, so the break was invisible there):
build-test-cpu and both sanitize-cpu jobs fail compiling
tests/vllm/models/test_glm_moe_dsa_schedule.cpp:304: cannot convert 'vllm::mla::MlaSharedSelection*' to 'vt::Tensor*'.
- Both
windows-msvc-* jobs fail at
src/vllm/model_executor/models/glm5_next_weights.cpp(223,31): error C2220 (warning treated as error).
Attribution
Both facets block green CI for every open PR until repaired; neither is caused by any open PR's changes.
FOLLOWING_AGENTS_PROTOCOL
Following-Agents-Protocol: true
AI-Assisted: true
Assisted-by: AGENT:zai-glm-5.3-flash [maki]
Observed
Against main
da2291339(first executed on PR #2370's merge ref; main's own run had these jobs cancelled before executing, so the break was invisible there):build-test-cpuand bothsanitize-cpujobs fail compilingtests/vllm/models/test_glm_moe_dsa_schedule.cpp:304:cannot convert 'vllm::mla::MlaSharedSelection*' to 'vt::Tensor*'.windows-msvc-*jobs fail atsrc/vllm/model_executor/models/glm5_next_weights.cpp(223,31):error C2220(warning treated as error).Attribution
ee5c86031(MODEL-TEXT-GLM-MOE-DSA W4);11f34effb(the same row's W9, "pass W9's shared selection to the parameter it names") changed the signature without carrying the test. Owner: MODEL-TEXT-GLM-MOE-DSA.glm5_next_weights.cpp(223)last changed ina36add6a8(MODEL-MM-GLM53-FLASH fix(MODEL-MM-GLM53-FLASH): readattention.key_lengththe way llama.cpp writes it, and DERIVE the KDA head count #2278). The in-flightrow/BUILD-CPU-WERROR-*branches guard MXFP4 tests underVT_MARLIN_NVFP4and do not touch this file. Owner: MODEL-MM-GLM53-FLASH.Both facets block green CI for every open PR until repaired; neither is caused by any open PR's changes.
FOLLOWING_AGENTS_PROTOCOL
Following-Agents-Protocol: true
AI-Assisted: true
Assisted-by: AGENT:zai-glm-5.3-flash [maki]