Back to Newsroom

Liquid AI Releases LFM2.5-2.6B: An On-Device Agentic Model With 128K Context, Tool Calling, And Open Weights

By Modelverse Editorial·August 7, 2026·2 min read
Liquid AI Releases LFM2.5-2.6B: An On-Device Agentic Model With 128K Context, Tool Calling, And Open Weights

Liquid AI's Latest Release: LFM2.5-2.6B

Liquid AI has unveiled LFM2.5-2.6B, a groundbreaking on-device agentic model designed to operate on a wide range of devices, from phones and laptops to PCs and robots. This innovative model boasts an impressive 2.69B total parameters, a vast 131,072-token context window, and a substantial 128,000-token vocabulary. By leveraging pre-training with approximately 34 trillion tokens, LFM2.5-2.6B demonstrates exceptional capabilities in planning, tool calling, and multi-step task execution.

Technical Capabilities and Open Access

The model's architecture consists of 30 layers, featuring a combination of double-gated short convolution blocks and grouped-query attention blocks. Notably, Liquid AI has made both checkpoints publicly available on Hugging Face under the lfm1.0 license, allowing developers to access the model's weights in various formats. This open approach enables seamless integration with popular frameworks, including llama.cpp, vLLM, SGLang, and LM Studio. By keeping inference local, LFM2.5-2.6B ensures that data remains on the device, minimizing marginal costs and enhancing user privacy.

Implications for Developers and Researchers

The release of LFM2.5-2.6B is significant, as it offers competitive performance with larger models while maintaining a relatively smaller size. Liquid AI's model has shown impressive results in instruction-following and tool-use benchmarks, outperforming larger models in several areas. This development has far-reaching implications for developers and researchers, as it enables the creation of more efficient, privacy-focused, and powerful AI applications that can operate effectively on a wide range of devices.

ai-newsbreakingmarktechpost

Footnotes & Primary References

Related content

Writer introduces new AI model and upgraded harness to contain token costs

Built as a post-training variation on Z.ai's open source model GLM-5.2, Writer says the new system should provide deployment-ready capabilities at a much lower price.

Read article

OpenAI introduces 'Ultrafast,' a new mode that makes GPT-5.6 Sol work at 14x the speed

OpenAI is launching a preview of a sped up version of its latest, most powerful model, in an effort to court enterprise users.

Read article

Nvidia's new $500B plan is risky but brilliant, especially for aging GPUs

Nvidia has a plan to make sure its GPUs won't lose value. It wants to convince a new crop of financiers to keep lending for AI buildouts.

Read article