Optimal settings for 2 GPUs in LM Studio
A user is seeking optimal settings for running large language models in LM Studio with a dual-GPU setup, comprising a 5070ti and a 5060ti, each with 16GB VRAM, totaling 32GB. They are attempting to run 16-18GB models like Qwen, Cydonia, and Skyfall, often with context lengths of 16384 or 32768. LM Studio indicates these models would use 17GB of VRAM, which is within the reported 32GB total. Despite seemingly sufficient VRAM, model loading frequently fails.
This post details a specific dual-GPU setup with 32GB VRAM, unlike general inquiries, and highlights a common issue where LM Studio reports sufficient VRAM but models fail to load.
Time & source
Times shown in UTC
Display time zone: UTC
Local time zone unavailable; showing UTC.
IngestedOffset at this time: UTC+0Oct 4, 2026, 14:00 UTC
- Ingested
- Oct 4, 2026, 14:00
- Source type
- Dev community
Full text isn't available here.
Read at source →