Skip to content

Releases: ZihanLiao/llama.cpp

b2001

Choose a tag to compare

@github-actions github-actions released this 29 Jan 09:41
fbe7dfa
ggml : add max buffer sizes to opencl and metal backends (#5181)

b1883

Choose a tag to compare

@github-actions github-actions released this 16 Jan 13:06
122ed48
examples : fix and improv docs for the grammar generator (#4909)

* Create pydantic-models-to-grammar.py

* Added some comments for usage

* Refactored Grammar Generator

Added example and usage instruction.

* Update pydantic_models_to_grammar.py

* Update pydantic-models-to-grammar-examples.py

* Renamed module and imported it.

* Update pydantic-models-to-grammar.py

* Renamed file and fixed grammar generator issue.

* Fixed some issues and bugs of the grammar generator. Imporved Documentation

* Update pydantic_models_to_grammar.py

b1818

Choose a tag to compare

@github-actions github-actions released this 11 Jan 09:43
64802ec
sync : ggml

b1787

Choose a tag to compare

@github-actions github-actions released this 08 Jan 01:50
b7e7982
readme : add lgrammel/modelfusion JS/TS client for llama.cpp (#4814)

b1728

Choose a tag to compare

@github-actions github-actions released this 30 Dec 02:24
a20f3c7
CUDA: fix tensor core logic for Pascal and HIP (#4682)