Difference in different quantization methods #2094
Answered
by
Green-Sky
sussyboiiii
asked this question in
Q&A
|
Hello, |
Answered by
Green-Sky
Jul 4, 2023
Replies: 4 comments 10 replies
|
The ppl column is perplexity increase relative to unquantized. |
7 replies
|
Брат, как сделать так, чтобы она информацию из реального времени брала? |
1 reply
|
Why q8, f16 and f32 are not recommended even if there is low quality loss. |
1 reply
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment



K-quantizations should be better, at the same file size, then the other ones. S M L means small medium large :)
more details can be found here: #1684