跳到正文
HNHacker News·
暂不在当前实时榜单

AI is now capable of developing its own inference hardware

AI 摘要

An open-source AI accelerator has been developed by AI, demonstrating its capability in creating inference hardware. Performance metrics show various models like LFM2.5-230M, Qwen3-0.6B, and Gemma 4 E2B achieving different token per second rates and DRAM usage. For instance, LFM2.5-230M int8 reached 59.0 tok/s with 14.5 GB/s DRAM. The accelerator also features faster prefill, though it is still limited by the matrix unit's multiply rate.

为什么是这条

This report details the first instance of an AI developing its own inference hardware, unlike previous AI applications that focused on software or design.

时间与来源

时间显示为 UTC

显示时区:UTC

本地时区尚不可用,暂时显示 UTC。

收录当时偏移:UTC+02026年10月6日 17:00 UTC

收录
2026年10月6日 17:00
来源类型
开发者社区

本站未收录正文。

前往源站阅读 →
来源·Hacker News·github.com