The AI Front Page

Reading signals from this article are folded back into your front page ranking on this device.

Agents/Hugging Face Blog/August 4, 2026 at 1:58 PM

Deploy local agents everywhere with LFM2.5-2.6B

Deploy local agents everywhere with LFM2.5-2.6B

Agents / Hugging Face Blog
Source

Follow Hugging Face Blog to make it a durable For You signal.

Liquid AI has released LFM2.5-2.6B, a 2.6 billion parameter model specifically designed for running agentic AI workloads entirely on-device. It enables private, cost-free deployment of capable agents on laptops, phones, and other everyday hardware, matching or exceeding the performance of models up to 4× its size on tool use, instruction following, and multi-step agent tasks. The model achieves 220 tokens/s on an Apple M5 Max and 113 tokens/s on an AMD Ryzen AI Max+ CPU, and it is available under open weights on Hugging Face with day-one support across common inference frameworks (llama.cpp, MLX, vLLM, SGLang, ONNX). This makes high-volume, locally run agentic applications practical without cloud dependencies.

Deploy local agents everywhere with LFM2.5-2.6B | The AI Front Page