v13
- The app allows benchmarking the model used in a chat. Go to
Settings > Benchmark Modelin any
chat and click 'Start Benchmarking'. The benchmark results arepp(tokens/sec for prompt
processing) andtg(tokens/sec for token generation).
Settings > Benchmark Model in anypp (tokens/sec for prompttg (tokens/sec for token generation).