v0.0.4-gptq-llama-triton
·
10 commits
to main
since this release
- Separate models in their own sub directories to prevent overriding configs when changing models
- Add triton support