learn less is more
try to use less code to get insight of different frameworks.
Hello world
- neon(TODO)
-
CUDA kernel
-
CUDA kernel in python
-
CUDA kernel in tensorrt(TODO))
-
opencl
-
cublas
- self attention
- multi-head attention
- cross attention
jit(TODO)
torchscrpt(TODO)
- onnx export
- onnx graphsurgeon