·
Archived topic · 归档话题,来源已停止追踪
LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge
LFM2.5-VL-3B is a vision-language model designed for on-device and real-time applications, offering enhanced vision capabilities for edge computing. This model, which understands documents and screens, grounds objects, and can call tools, provides direct answers rather than reasoning to maintain speed. Benchmarks show LFM2.5-VL-3B (3.1B) achieving an average score of 69.4, outperforming LFM2-VL-3B (3.1B) at 57.2 and gemma-4-E2B-it (5.1B) at 52.0 across various tasks including MMStar, MME, RealWorldQA, and OCRBench v2 (En).
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年8月12日 15:00 UTC
- 收录
- 2026年8月12日 15:00
- 来源类型
- 未分类