What's new
- Expandable thinking stream: click to expand and watch reasoning tokens stream in live while model is thinking
- Auto routing load spread: concurrent chat requests spread across top-scoring models instead of queueing on one
- CDN-ready cache headers: hashed static assets served with immutable cache headers, SSE unbuffered for Cloudflare/CDN compatibility
- Updated project description and chat placeholder
Install (macOS Apple Silicon)
curl -fsSL https://github.com/michaelneale/decentralized-inference/releases/latest/download/mesh-llm-aarch64-apple-darwin.tar.gz | tar xz && sudo mv mesh-bundle/* /usr/local/bin/