Skip to content
RCreddit.com·
Not on the current live radar

Got an old slow low vram GPU laying around? Might be worth it to use for Just Vision mmproj llama.cpp

AI summary

Users with older, low VRAM GPUs can repurpose them for mmproj in llama.cpp to significantly improve performance. Instead of using --no-mmproj-offload, which is slow, especially for agentic coding, a secondary GPU can be dedicated to mmproj using --mmdev CUDA1. This setup offers a magnitude faster processing for mmproj without impacting the primary GPU's inference speed, making efficient use of otherwise underutilized hardware.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

IngestedOffset at this time: UTC+0Sep 12, 2026, 06:01 UTC

Ingested
Sep 12, 2026, 06:01
Source type
Dev community
Article

Full text isn't available here.

Read at source →
Source·reddit.com