π§ FastFlowLM v0.9.25 β New Models + OpenAI API Fixes
We're excited to introduce FastFlowLM v0.9.25, marking a key milestone with the integration of the new LFM2.5 model, freshly unveiled at CES 2026 (Jan 5th). This release also includes improvements to API compatibility and instruction-style models.
π New Model Support
-
LFM2.5-1.2B-Instruct
πΈ Debuted at CES2026
The newest addition to the LFM family, tuned for instruction-following. It features improved responsiveness and latency, ideal for interactive applications on AMD NPU. -
Phi4-mini-instruct
A compact instruction model tailored for devices with limited memory β great for summarization and low-resource tasks.
π οΈ Fixes & Improvements
- β
Fixed bugs related to generation parameters (
top_k,top_p, etc.) not being respected in OpenAI-compatible REST APIs. - Ensures correct behavior when adjusting generation strategy through API calls.
π With this update, FastFlowLM continues its mission to support the latest LLMs and provide an efficient, private, and developer-friendly experience on AMD Ryzen AI NPUs.