Deploy local agents everywhere with LFM2.5-2.6B

Huggingface··Submitted by Mads Kristian Nylund
AI DevelopmentAI ToolsAI Infrastructure

LFM2.5-2.6B is a lightweight model designed for on-device agent deployment, offering efficient inference on various hardware platforms. It excels in instruction-following and tool-use benchmarks, outperforming smaller models while maintaining strong performance in knowledge and math tasks. The model achieves high inference speeds, such as 220 tokens per second on an Apple M5 Max and 113 tokens per second on an AMD Ryzen CPU, with low memory usage, making it suitable for mobile and portable devices.

Read Article

More from Huggingface

Related Articles