We're excited to introduce FastFlowLM v0.9.25, marking a key milestone with the integration of the new LFM2.5 model, freshly unveiled at CES 2026 (Jan 5th). This release also includes improvements to API compatibility and instruction-style models.
🚀 New Model Support
-
LFM2.5-1.2B-Instruct
🔸 Debuted at CES2026
The newest addition to the LFM family, tuned for instruction-following. It features improved responsiveness and latency, ideal for interactive applications on AMD NPU. -
Phi4-mini-instruct
A compact instruction model tailored for devices with limited memory — great for summarization and low-resource tasks.
🛠️ Fixes & Improvements
- ✅ Fixed bugs related to generation parameters (
top_k,top_p, etc.) not being respected in OpenAI-compatible REST APIs. - Ensures correct behavior when adjusting generation strategy through API calls.
🎉 With this update, FastFlowLM continues its mission to support the latest LLMs and provide an efficient, private, and developer-friendly experience on AMD Ryzen AI NPUs.