Specifications
MoE Efficiency
24B quality, 2B inference cost
Laptop-Ready
Runs on laptops and single GPUs
Tool Calling
Native function calling support
Quick Start
- Transformers
- llama.cpp
- vLLM
Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
🚀 New: LFM2.5-VL-450M — our smallest vision model is now available! Learn more →
24B parameter Mixture-of-Experts model with 2B active parameters — our largest model for laptops and single-GPU applications
| Property | Value |
|---|---|
| Parameters | 24B (2B active) |
| Context Length | 32K tokens |
| Architecture | LFM2 (MoE) |
Was this page helpful?