跳到正文
RCreddit.com·
暂不在当前实时榜单

My foray into local ai. Two BC-250 ex mining apus running Qwen3.6-35B-A3B Q4_K_M at 60 tok/s with 64k context

AI 摘要

A user is running local AI with two BC-250 ex-mining APUs, costing $115 each, connected via llama.cpp with Vulkan and RPC on Bazzite. These boards offer approximately 27GB of combined GPU memory and communicate over 1gb Ethernet. The setup, totaling around $300 including the PSU, achieves 60 tok/s with 64k context when running Qwen3.6-35B-A3B Q4_K_M. The user plans to expand to six boards to test Qwen 3.8 flash.

为什么是这条

This report details a specific, low-cost hardware setup for local AI, unlike general discussions of AI models or software, and includes concrete performance metrics for Qwen3.6-35B-A3B.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月24日 12:02 UTC

收录
2026年9月24日 12:02
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·reddit.com