Skip to content

v0.1.9

Choose a tag to compare

@rybruscoe rybruscoe released this 05 Sep 07:35
· 29 commits to main since this release

Exact profile on CUDA: llama-cpp-et's CUDA backend carries the exact-quire b-posit8 kernel (bit-identical to the CPU kernel; every matmul re-executed by the Python and Go verifiers). INVAR pins the compute device and offloaded layer count in the receipt. The block-scale rule is now integer-exact in all implementations. New: invar-statement (Go COSE_Sign1 verifier), docs/ARCHITECTURE.md.