Specifications
Edge Deployment
Optimized for resource-constrained devices
Low Latency
Fast inference for real-time applications
Fine-tunable
TRL compatible (SFT, DPO, GRPO)
Quick Start
- Transformers
- llama.cpp
- vLLM
- SGLang
Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
🚀 New: LFM2.5-VL-450M — our smallest vision model is now available! Learn more →
Mid-sized 700M parameter model for deploying on most devices
| Property | Value |
|---|---|
| Parameters | 700M |
| Context Length | 32K tokens |
| Architecture | LFM2 (Dense) |
Was this page helpful?