跳到正文
RCreddit.com·

Do you need some extra memory on your DGX Spark?

AI 摘要

A new repository helps DGX Spark users expand memory by offloading the spec-decode draft model to a spare 10-24 GB GPU. This frees up gigabytes of memory on the Spark, allowing for increased context or improved quantization quality. The solution supports both TCP and RDMA, and is shipped as eugr-vllm compatible modifications, available at the provided GitHub link.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

发布当时偏移:UTC+02026年9月29日 05:13 UTC

收录当时偏移:UTC+02026年9月29日 20:00 UTC

发布
2026年9月29日 05:13
收录
2026年9月29日 20:00
来源类型
开发者社区
档位
社区
信源状态
正常

档位是按信源手工设定的编辑判断,不是逐条打分。

I created this repo to help the DGX Spark users that have a spare 10-24 GB GPU at home to squeeze some extra memory out of a single Spark or a Sparks cluster.

It moves the spec-decode draft model off your Sparks onto that GPU: the freed GB of memory can be used for extra context, or better quant quality. Supports both TCP and RDMA, shipped as eugr-vllm compatible mods:

来源·reddit.com