This repository was archived by the owner on Sep 28, 2025. It is now read-only.
Repository navigation
v0.1.6
Changed:
- Updated huggingface-hub.
Fixed:
- llama.__init__ now correctly imports submodules and handles CPU and CUDA backends.
- OpenAI: ctx_size: int = config.max_position_embeddings if max_tokens is None else max_tokens.