跳到正文
RCreddit.com·
暂不在当前实时榜单

apple/LensVLM-9B · Hugging Face

AI 摘要

LensVLM is a 9B Vision Language Model (VLM) designed to process compressed images of text. It selectively expands only the relevant pages to their uncompressed form using learned tools. The model, available on Hugging Face as bartowski/LensVLM-9B-GGUF, has its accompanying source code distributed separately under the Apple Sample Code License.

为什么是这条

This 9B Vision Language Model from Apple is the first to use selective expansion of compressed text images, unlike prior models that process all data uniformly.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月23日 23:01 UTC

收录
2026年9月23日 23:01
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com