Specifications
MoE Efficiency
8B quality, 1.5B inference cost
On-Device
Runs on phones and laptops
Tool Calling
Native function calling support
Quick Start
- Transformers
- llama.cpp
- vLLM
- SGLang
Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
🚀 New: LFM2.5-VL-450M — our smallest vision model is now available! Learn more →
8B parameter Mixture-of-Experts model with 1.5B active parameters for fast, high-quality inference
| Property | Value |
|---|---|
| Parameters | 8B (1.5B active) |
| Context Length | 32K tokens |
| Architecture | LFM2 (MoE) |
Was this page helpful?