This repository was archived by the owner on Apr 19, 2024. It is now read-only.
1.0.1-cpu
This release uses CPU inference with llama.cpp, with optional GPU layer offloading.
This release uses CPU inference with llama.cpp, with optional GPU layer offloading.