Skip to content

llama cpp and gguf

github-actions[bot] edited this page Sep 30, 2026 · 2 revisions

llama.cpp and GGUF

The jevos-v2 release ships the model as GGUF files for llama.cpp and similar tools; jev itself runs 8-bit OpenVINO weights and uses llama.cpp only as its tokenizer. These pages explain the pieces for someone deploying a classifier rather than chatting with a model.

Guides

Measurements

Comparisons

Speed

Probability and thresholds

Question design

Use cases

Evaluation

Agents and routing

Integrations

Local and private AI

llama.cpp and GGUF

Clone this wiki locally