Qwen3.5/3.6 dense with MTP supported.
Flash-Linear-Attention-NPU was leveraged to accelerate the training of Qwen3.5 and Qwen3.6 on NPUs, achieving a 3–10x speedup.
Tensor parallelism with Megatron and MindSpeed will no longer be required as dependencies on the NPU platform.
Upgrade the NPU platform environment to Python 3.11.