Deploy local agents everywhere with LFM2.5-2.6B
Deploy local agents everywhere with LFM2.5-2.6B
Follow Hugging Face Blog to make it a durable For You signal.
Liquid AI has released LFM2.5-2.6B, a 2.6 billion parameter model specifically designed for running agentic AI workloads entirely on-device. It enables private, cost-free deployment of capable agents on laptops, phones, and other everyday hardware, matching or exceeding the performance of models up to 4× its size on tool use, instruction following, and multi-step agent tasks. The model achieves 220 tokens/s on an Apple M5 Max and 113 tokens/s on an AMD Ryzen AI Max+ CPU, and it is available under open weights on Hugging Face with day-one support across common inference frameworks (llama.cpp, MLX, vLLM, SGLang, ONNX). This makes high-volume, locally run agentic applications practical without cloud dependencies.