What's new
- Expandable thinking stream: while a model is thinking, click to expand and watch reasoning tokens stream in live (previously just showed a static spinner)
- Auto routing load spread: concurrent chat requests in auto mode now spread across top-scoring models instead of always queueing on the same one
- Updated project description and chat placeholder
Install (macOS Apple Silicon)
curl -fsSL https://github.com/michaelneale/decentralized-inference/releases/latest/download/mesh-llm-aarch64-apple-darwin.tar.gz | tar xz && sudo mv mesh-bundle/* /usr/local/bin/