Skip to content

Neko.js v1.0

Choose a tag to compare

@YueyuHoshizora YueyuHoshizora released this 02 Oct 12:08
· 64 commits to main since this release

繁體中文

Neko.js v1.0 是以固定 Qwen 模型進行本機圖片+提示推理的原型。Node.js 使用原生 ONNX Runtime 1.30.0;瀏覽器使用 WebGPU 與 ONNX Runtime Web 1.26.0-dev.20260416-b7804b056c。另使用 Transformers.js 4.2.0、Qwen3.5-0.8B Q4 embedding/decoder 與 FP16 vision encoder。已驗證 13 個模型檔的完整性及快取離線重新載入;另提供 inert 網頁擷取、來源資訊與 Markdown 報告 helper。

這不是完整的「輸入網址並產生結構化報告」產品。瀏覽器 CPU/WASM 對此模型不支援(GatherBlockQuantized(1));Node CPU 尚未驗證。不保證自動回退,也不宣稱所有運算子皆在 GPU 執行。尚未提供相同原始圖片與提示的 Ollama 執行結果作為參考,故無法宣稱 parity。請參閱中英文使用指南了解已驗證行為及限制。

English

Neko.js v1.0 is a local image-plus-prompt inference prototype using the pinned Qwen model. Node.js uses native ONNX Runtime 1.30.0; the browser uses WebGPU with ONNX Runtime Web 1.26.0-dev.20260416-b7804b056c. It also uses Transformers.js 4.2.0, Qwen3.5-0.8B Q4 embedding/decoder, and an FP16 vision encoder. Integrity of 13 model files and cached offline reload were verified. It also includes inert web extraction, provenance, and Markdown report helpers.

This is not a complete URL-to-structured-report product. Browser CPU/WASM is unsupported for this model (GatherBlockQuantized(1)); Node CPU is unverified. Automatic fallback is not claimed, nor is execution of every operator on GPU. No Ollama run using the same original image and prompt was provided as a reference, so parity is not claimed. See the bilingual usage guides for verified behavior and limitations.