Skip to content
RCreddit.com·
Not on the current live radar

Infermeld: a Linux kit for running one GGUF across AMD + NVIDIA GPUs with llama.cpp

AI summary

Infermeld is an open-source Linux kit designed to run a single GGUF model across both AMD and NVIDIA GPUs using llama.cpp. It was tested with an AMD RX 6900 XT (16GB) and an NVIDIA RTX 3080 (10GB) using the Qwen3.6-35B-A3B, UD-Q4_K_M GGUF model. The kit utilizes Vulkan and CUDA backends in split mode, reserving 8,192 tokens for context, and supports loading and short completions with MTP both off and on.

Why this one

This kit is among the first to enable running a single GGUF model across both AMD and NVIDIA GPUs simultaneously, unlike previous solutions that typically target one vendor.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

IngestedOffset at this time: UTC+0Oct 5, 2026, 11:00 UTC

Ingested
Oct 5, 2026, 11:00
Source type
Dev community

Full text isn't available here.

Read at source →
Source·reddit.com