·
Archived topic · source no longer tracked
LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge
LFM2.5-VL-3B is a vision-language model designed for on-device and real-time applications, offering enhanced vision capabilities for edge computing. This model, which understands documents and screens, grounds objects, and can call tools, provides direct answers rather than reasoning to maintain speed. Benchmarks show LFM2.5-VL-3B (3.1B) achieving an average score of 69.4, outperforming LFM2-VL-3B (3.1B) at 57.2 and gemma-4-E2B-it (5.1B) at 52.0 across various tasks including MMStar, MME, RealWorldQA, and OCRBench v2 (En).
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
IngestedOffset at this time: UTC+0Aug 12, 2026, 15:00 UTC
- Ingested
- Aug 12, 2026, 15:00
- Source type
- Unclassified