跳到正文
RCreddit.com·
暂不在当前实时榜单

DeepSeek-V4-Flash-Vision-Exp (285B MoE) on 10-12x RTX 3090 — spec decoding, vision

AI 摘要

The DeepSeek-V4-Flash-Vision-Exp (285B MoE) model has been successfully run on 10-12x RTX 3090 GPUs, achieving 60+ tok/s decode on 10 GPUs and 120+ tok/s on 12 GPUs. This setup supports vision, speculative decoding, and tool calls, with a 1M context without offload or 4M with RAM offload. The implementation uses FP4 experts and FP8 attention, with 157 GB weights, and is fully documented and reproducible via a Docker image and GitHub repository.

为什么是这条

This is the first public report of DeepSeek-V4-Flash-Vision-Exp running on consumer-grade Ampere GPUs, unlike previous benchmarks on more powerful hardware.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年9月9日 13:00 UTC

收录
2026年9月9日 13:00
来源类型
开发者社区
正文

本站未收录正文。

前往源站阅读 →
来源·reddit.com