Skip to content

v1.3.0 — Neue Modell-Generation 2025/2026

Latest

Choose a tag to compare

@ogerly ogerly released this 27 Sep 15:12
· 3 commits to main since this release

🏟️ WebGPU-Arena v1.3.0

Die neue Modell-Generation ist da. Die Arena fährt jetzt die 2025/2026er-Klasse – die komplette 2024er-Reihe ist archiviert (ELO-Historie bleibt erhalten).

✨ Neue Arena-Modelle

  • Qwen3.5 4B – Flaggschiff (2026): Vision-fähig, 262k Kontext, GPQA Diamond 76.2 (über GPT-OSS-20B: 71.5)
  • Phi-4 mini (3.8B) – Reasoning-Spezialist
  • Ministral 3 3B – Instruct + Reasoning-Variante (Chain-of-Thought)
  • Hermes 3 3B – Nachfolger des Llama-3.2-Champions
  • Qwen3.5 2B, Qwen3 1.7B (Thinking-Modus), Qwen 2.5 3B, Gemma 3 1B

🗄️ Archiv

Llama 3.2 (1B/3B), Gemma 2 2B, SmolLM2 1.7B, Qwen 2.5 (0.5B/1.5B) und TinyLlama sind ausgemustert: sichtbar als «Archiv» in der Bibliothek und mit «Retired»-Tag in der Rangliste.

🛠️ Weitere Änderungen

  • WebGPU Benchmark (/benchmark): MatMul / SAXPY / GEMM – GPU vs. CPU (WP-015)
  • Fixes: Arena-Selektor ohne Bild-Modelle, Leaderboard-NaN-Sort-Bug, vite-plugin-pwa 1.3.0 (Vite-8-Peer-Konflikt behoben)
  • Deps: @mlc-ai/web-llm 0.2.85

📚 Doku

README, CHANGELOG (inkl. nachgetragener v1.2.0), WP-018 Workpaper, LTM-Index und Diary aktualisiert.


100% lokal, ohne API-Keys, WebGPU-powered. AAMS-dokumentiert.