Skip to content

[Bug]: [registry.py:410] ValueError: infer_schema(func): Parameter block_size has unsupported type list[int]. #21991

Description

@hnhyzz

Your current environment

The output of python collect_env.py
==============================
        System Info
==============================
OS                           : Ubuntu 24.04.2 LTS (x86_64)
GCC version                  : (Ubuntu 13.3.0-6ubuntu2~24.04) 13.3.0
Clang version                : 19.0.0git (https://github.com/RadeonOpenCompute/llvm-project roc-6.4.0 25133 c7fe45cf4b819c5991fe208aaa96edf142730f1d)
CMake version                : version 3.31.2
Libc version                 : glibc-2.39

==============================
       PyTorch Info
==============================
PyTorch version              : 2.6.0+git45896ac
Is debug build               : False
CUDA used to build PyTorch   : N/A
ROCM used to build PyTorch   : 6.4.43482-0f2d60242

==============================
      Python Environment
==============================
Python version               : 3.12.9 | packaged by Anaconda, Inc. | (main, Feb  6 2025, 18:56:27) [GCC 11.2.0] (64-bit runtime)
Python platform              : Linux-6.8.0-60-generic-x86_64-with-glibc2.39

==============================
       CUDA / GPU Info
==============================
Is CUDA available            : True
CUDA runtime version         : Could not collect
CUDA_MODULE_LOADING set to   : LAZY
GPU models and configuration : AMD Instinct MI100 (gfx908:sramecc+:xnack-)
Nvidia driver version        : Could not collect
cuDNN version                : Could not collect
HIP runtime version          : 6.4.43482
MIOpen runtime version       : 3.4.0
Is XNNPACK available         : True

==============================
          CPU Info
==============================
Architecture:                         x86_64
CPU op-mode(s):                       32-bit, 64-bit
Address sizes:                        43 bits physical, 48 bits virtual
Byte Order:                           Little Endian
CPU(s):                               32
On-line CPU(s) list:                  0-31
Vendor ID:                            AuthenticAMD
BIOS Vendor ID:                       Advanced Micro Devices, Inc.
Model name:                           AMD EPYC 7282 16-Core Processor
BIOS Model name:                      AMD EPYC 7282 16-Core Processor                 Unknown CPU @ 2.8GHz
BIOS CPU family:                      107
CPU family:                           23
Model:                                49
Thread(s) per core:                   1
Core(s) per socket:                   16
Socket(s):                            2
Stepping:                             0
Frequency boost:                      disabled
CPU(s) scaling MHz:                   65%
CPU max MHz:                          2800.0000
CPU min MHz:                          1500.0000
BogoMIPS:                             5599.51
Flags:                                fpu vme de pse tsc msr pae mce cx8 apic sep mtrr pge mca cmov pat pse36 clflush mmx fxsr sse sse2 ht syscall nx mmxext fxsr_opt pdpe1gb rdtscp lm constant_tsc rep_good nopl nonstop_tsc cpuid extd_apicid aperfmperf rapl pni pclmulqdq monitor ssse3 fma cx16 sse4_1 sse4_2 movbe popcnt aes xsave avx f16c rdrand lahf_lm cmp_legacy svm extapic cr8_legacy abm sse4a misalignsse 3dnowprefetch osvw ibs skinit wdt tce topoext perfctr_core perfctr_nb bpext perfctr_llc mwaitx cpb cat_l3 cdp_l3 hw_pstate ssbd mba ibrs ibpb stibp vmmcall fsgsbase bmi1 avx2 smep bmi2 cqm rdt_a rdseed adx smap clflushopt clwb sha_ni xsaveopt xsavec xgetbv1 xsaves cqm_llc cqm_occup_llc cqm_mbm_total cqm_mbm_local clzero irperf xsaveerptr rdpru wbnoinvd amd_ppin arat npt lbrv svm_lock nrip_save tsc_scale vmcb_clean flushbyasid decodeassists pausefilter pfthreshold avic v_vmsave_vmload vgif v_spec_ctrl umip rdpid overflow_recov succor smca sev sev_es
Virtualization:                       AMD-V
L1d cache:                            1 MiB (32 instances)
L1i cache:                            1 MiB (32 instances)
L2 cache:                             16 MiB (32 instances)
L3 cache:                             128 MiB (8 instances)
NUMA node(s):                         2
NUMA node0 CPU(s):                    0-15
NUMA node1 CPU(s):                    16-31
Vulnerability Gather data sampling:   Not affected
Vulnerability Itlb multihit:          Not affected
Vulnerability L1tf:                   Not affected
Vulnerability Mds:                    Not affected
Vulnerability Meltdown:               Not affected
Vulnerability Mmio stale data:        Not affected
Vulnerability Reg file data sampling: Not affected
Vulnerability Retbleed:               Mitigation; untrained return thunk; SMT disabled
Vulnerability Spec rstack overflow:   Mitigation; SMT disabled
Vulnerability Spec store bypass:      Mitigation; Speculative Store Bypass disabled via prctl
Vulnerability Spectre v1:             Mitigation; usercopy/swapgs barriers and __user pointer sanitization
Vulnerability Spectre v2:             Mitigation; Retpolines; IBPB conditional; STIBP disabled; RSB filling; PBRSB-eIBRS Not affected; BHI Not affected
Vulnerability Srbds:                  Not affected
Vulnerability Tsx async abort:        Not affected

==============================
Versions of relevant libraries
==============================
[pip3] conch-triton-kernels==1.2.1
[pip3] mypy==1.13.0
[pip3] mypy-extensions==1.0.0
[pip3] numpy==1.26.4
[pip3] onnx==1.17.0
[pip3] onnxscript==0.1.0.dev20240817
[pip3] optree==0.13.0
[pip3] pyzmq==27.0.0
[pip3] torch==2.6.0+git45896ac
[pip3] torchvision==0.21.0+4040d51
[pip3] transformers==4.54.1
[pip3] triton==3.2.0
[conda] No relevant packages

==============================
         vLLM Info
==============================
ROCM Version                 : 6.4.43482-0f2d60242
Neuron SDK Version           : N/A
vLLM Version                 : 0.10.1.dev235+g055bd3978 (git sha: 055bd3978)
vLLM Build Flags:
  CUDA Archs: Not Set; ROCm: Disabled; Neuron: Disabled
GPU Topology:
  ============================ ROCm System Management Interface ============================
================================ Weight between two GPUs =================================
       GPU0         GPU1         GPU2         GPU3         GPU4         GPU5         GPU6         GPU7         
GPU0   0            40           15           15           72           72           72           15           
GPU1   40           0            40           40           15           15           15           72           
GPU2   15           40           0            15           72           72           72           15           
GPU3   15           40           15           0            72           72           72           15           
GPU4   72           15           72           72           0            15           15           40           
GPU5   72           15           72           72           15           0            15           40           
GPU6   72           15           72           72           15           15           0            40           
GPU7   15           72           15           15           40           40           40           0            

================================= Hops between two GPUs ==================================
       GPU0         GPU1         GPU2         GPU3         GPU4         GPU5         GPU6         GPU7         
GPU0   0            2            1            1            3            3            3            1            
GPU1   2            0            2            2            1            1            1            3            
GPU2   1            2            0            1            3            3            3            1            
GPU3   1            2            1            0            3            3            3            1            
GPU4   3            1            3            3            0            1            1            2            
GPU5   3            1            3            3            1            0            1            2            
GPU6   3            1            3            3            1            1            0            2            
GPU7   1            3            1            1            2            2            2            0            

=============================== Link Type between two GPUs ===============================
       GPU0         GPU1         GPU2         GPU3         GPU4         GPU5         GPU6         GPU7         
GPU0   0            PCIE         XGMI         XGMI         PCIE         PCIE         PCIE         XGMI         
GPU1   PCIE         0            PCIE         PCIE         XGMI         XGMI         XGMI         PCIE         
GPU2   XGMI         PCIE         0            XGMI         PCIE         PCIE         PCIE         XGMI         
GPU3   XGMI         PCIE         XGMI         0            PCIE         PCIE         PCIE         XGMI         
GPU4   PCIE         XGMI         PCIE         PCIE         0            XGMI         XGMI         PCIE         
GPU5   PCIE         XGMI         PCIE         PCIE         XGMI         0            XGMI         PCIE         
GPU6   PCIE         XGMI         PCIE         PCIE         XGMI         XGMI         0            PCIE         
GPU7   XGMI         PCIE         XGMI         XGMI         PCIE         PCIE         PCIE         0            

======================================= Numa Nodes =======================================
GPU[0]		: (Topology) Numa Node: 0
GPU[0]		: (Topology) Numa Affinity: 0
GPU[1]		: (Topology) Numa Node: 0
GPU[1]		: (Topology) Numa Affinity: 0
GPU[2]		: (Topology) Numa Node: 0
GPU[2]		: (Topology) Numa Affinity: 0
GPU[3]		: (Topology) Numa Node: 0
GPU[3]		: (Topology) Numa Affinity: 0
GPU[4]		: (Topology) Numa Node: 1
GPU[4]		: (Topology) Numa Affinity: 1
GPU[5]		: (Topology) Numa Node: 1
GPU[5]		: (Topology) Numa Affinity: 1
GPU[6]		: (Topology) Numa Node: 1
GPU[6]		: (Topology) Numa Affinity: 1
GPU[7]		: (Topology) Numa Node: 1
GPU[7]		: (Topology) Numa Affinity: 1
================================== End of ROCm SMI Log ===================================

==============================
     Environment Variables
==============================
PYTORCH_TESTING_DEVICE_ONLY_FOR=cuda
PYTORCH_TEST_WITH_ROCM=1
CUDA_VISIBLE_DEVICES=1,4,5,6
CUDA_VISIBLE_DEVICES=1,4,5,6
PYTORCH_ROCM_ARCH=gfx908
MAX_JOBS=32
LD_LIBRARY_PATH=/opt/ompi/lib:/opt/rocm/lib:/usr/local/lib:
NCCL_CUMEM_ENABLE=0
PYTORCH_NVML_BASED_CUDA_CHECK=1
TORCHINDUCTOR_COMPILE_THREADS=1
CUDA_MODULE_LOADING=LAZY


🐛 Describe the bug

When serving the Qwen3 model on the vllm=0.10.1 version, it reports the following error.

Command: vllm serve /models/Qwen3-8B/

INFO 07-31 05:30:20 [init.py:241] Automatically detected platform rocm.
�[1;36m(APIServer pid=670)�[0;0m INFO 07-31 05:30:40 [api_server.py:1774] vLLM API server version 0.10.1.dev235+g055bd3978
�[1;36m(APIServer pid=670)�[0;0m INFO 07-31 05:30:40 [utils.py:326] non-default args: {'model_tag': '/models/Qwen3-0.6B-GPTQ-Int8/', 'port': 8080, 'api_key': ['4families'], 'model': '/models/Qwen3-0.6B-GPTQ-Int8/', 'max_model_len': 8192, 'served_model_name': ['zz-model']}
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] Error in inspecting model architecture 'Qwen3ForCausalLM'
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] Traceback (most recent call last):
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 820, in _run_in_subprocess
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] returned.check_returncode()
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/subprocess.py", line 504, in check_returncode
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] raise CalledProcessError(self.returncode, self.args, self.stdout,
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] subprocess.CalledProcessError: Command '['/opt/conda/envs/py_3.12/bin/python3', '-m', 'vllm.model_executor.models.registry']' returned non-zero exit status 1.
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410]
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] The above exception was the direct cause of the following exception:
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410]
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] Traceback (most recent call last):
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 408, in _try_inspect_model_cls
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] return model.inspect_model_cls()
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ^^^^^^^^^^^^^^^^^^^^^^^^^
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 379, in inspect_model_cls
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] return _run_in_subprocess(
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ^^^^^^^^^^^^^^^^^^^
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 823, in _run_in_subprocess
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] raise RuntimeError(f"Error raised in subprocess:\n"
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] RuntimeError: Error raised in subprocess:
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] :128: RuntimeWarning: 'vllm.model_executor.models.registry' found in sys.modules after import of package 'vllm.model_executor.models', but prior to execution of 'vllm.model_executor.models.registry'; this may result in unpredictable behaviour
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] Traceback (most recent call last):
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 198, in _run_module_as_main
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 88, in _run_code
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 844, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] _run()
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 837, in _run
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] result = fn()
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ^^^^
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 380, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] lambda: _ModelInfo.from_model_cls(self.load_model_cls()))
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ^^^^^^^^^^^^^^^^^^^^^
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/registry.py", line 383, in load_model_cls
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] mod = importlib.import_module(self.module_name)
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/importlib/init.py", line 90, in import_module
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] return _bootstrap._gcd_import(name[level:], package, level)
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 1387, in _gcd_import
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 1360, in _find_and_load
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 1331, in _find_and_load_unlocked
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 935, in _load_unlocked
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 999, in exec_module
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "", line 488, in _call_with_frames_removed
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/qwen3.py", line 48, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from .qwen2 import Qwen2MLP as Qwen3MLP
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/models/qwen2.py", line 48, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from vllm.model_executor.model_loader.weight_utils import (
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/model_loader/init.py", line 11, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from vllm.model_executor.model_loader.bitsandbytes_loader import (
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/model_loader/bitsandbytes_loader.py", line 23, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from vllm.model_executor.layers.fused_moe import FusedMoE
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/layers/fused_moe/init.py", line 8, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from vllm.model_executor.layers.fused_moe.layer import (
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/layers/fused_moe/layer.py", line 26, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from vllm.model_executor.layers.fused_moe.modular_kernel import (
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/layers/fused_moe/modular_kernel.py", line 13, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from vllm.model_executor.layers.fused_moe.utils import ( # yapf: disable
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/layers/fused_moe/utils.py", line 9, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] from vllm.model_executor.layers.quantization.utils.fp8_utils import (
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/model_executor/layers/quantization/utils/fp8_utils.py", line 78, in
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] direct_register_custom_op(
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/vllm/utils/init.py", line 2522, in direct_register_custom_op
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] schema_str = torch.library.infer_schema(op_func,
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^^
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/torch/_library/infer_schema.py", line 106, in infer_schema
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] error_fn(
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] File "/opt/conda/envs/py_3.12/lib/python3.12/site-packages/torch/_library/infer_schema.py", line 58, in error_fn
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] raise ValueError(
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410] ValueError: infer_schema(func): Parameter block_size has unsupported type list[int]. The valid types are: dict_keys([<class 'torch.Tensor'>, typing.Optional[torch.Tensor], typing.Sequence[torch.Tensor], typing.List[torch.Tensor], typing.Sequence[typing.Optional[torch.Tensor]], typing.List[typing.Optional[torch.Tensor]], <class 'int'>, typing.Optional[int], typing.Sequence[int], typing.List[int], typing.Optional[typing.Sequence[int]], typing.Optional[typing.List[int]], <class 'float'>, typing.Optional[float], typing.Sequence[float], typing.List[float], typing.Optional[typing.Sequence[float]], typing.Optional[typing.List[float]], <class 'bool'>, typing.Optional[bool], typing.Sequence[bool], typing.List[bool], typing.Optional[typing.Sequence[bool]], typing.Optional[typing.List[bool]], <class 'str'>, typing.Optional[str], typing.Union[int, float, bool], typing.Union[int, float, bool, NoneType], typing.Sequence[typing.Union[int, float, bool]], typing.List[typing.Union[int, float, bool]], <class 'torch.dtype'>, typing.Optional[torch.dtype], <class 'torch.device'>, typing.Optional[torch.device]]). Got func with signature (A: torch.Tensor, B: torch.Tensor, As: torch.Tensor, Bs: torch.Tensor, block_size: list[int], output_dtype: torch.dtype = torch.float16) -> torch.Tensor)
�[1;36m(APIServer pid=670)�[0;0m ERROR 07-31 05:31:05 [registry.py:410]

Before submitting a new issue...

  • Make sure you already searched for relevant issues, and asked the chatbot living at the bottom right corner of the documentation page, which can answer lots of frequently asked questions.

Metadata

Metadata

Assignees

No one assigned

    Labels

    bugSomething isn't working

    Type

    No type

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions