github ROCm/FastFlowLM v0.9.24
πŸš€ FastFlowLM v0.9.24 β€” LiquidAI Upgrade, Prompt Caching, & Rock-Solid Stability

latest releases: v1.0.4, v1.0.3, v1.0.2...
8 months ago

FastFlowLM v0.9.24 is here with a powerful new LiquidAI model, smarter caching, safer downloads, and better runtime control β€” making your NPU experience faster, smoother, and more reliable than ever.


πŸ“š Expanded LiquidAI Support

Feature Details
New Model LFM2:2.6B added β€” delivers 31+ tokens/sec (tps) in decoding and 1000+ tps in prefill.
Faster Decoding LFM2:1.2B now reaches 63+ tokens/sec

πŸ” Redownload is required for the new LFM2:1.2B.


πŸ”‘ What’s New & Why It Matters

Feature Benefit
Prompt Cache (Server Mode) Reuses recent inputs to cut latency and speed up multi-turn conversations. Inspired by llamacpp.
Download Integrity Checks Every model download is now checksum-verified to ensure reliability and correctness. Huge thanks to @ramkrishna2910 for reporting and @jeremyfowers for guidance!
Interrupt During Decoding Stop generation mid-stream in serve mode for tighter control over long or unwanted outputs β€” another great suggestion from @jeremyfowers.

πŸ™Œ A huge, heartfelt thank you to our community and early adopters β€” your testing, feedback, bug reports, patience, and enthusiasm are what make FastFlowLM better every single day. We truly couldn’t do this without you.

βœ¨πŸŽ‰πŸŒ πŸŽ„
Happy Holidays! Wishing you warmth, joy, and inspiration β€” and we can’t wait to keep building amazing AI together in the new year.

Don't miss a new FastFlowLM release

NewReleases is sending notifications on new releases.