返回
RCreddit.com
16
·16小时前·开发者社区 · RSS

Is the US-China AI capability gap still meaningful for actual production workloads?

查看原文

热度趋势

趋势数据积累中

百分比基于当前可用热度信号,而非评论数或独立用户人数。

I've been using Chinese models more and more this year. Started with DeepSeek for reasoning stuff, moved to Qwen for longer context work, tried GLM when it had that mini DeepSeek moment on OpenRouter. At this point the rotation is mostly Chinese models with Claude as the fallback for tricky creative tasks.

This week I finally got around to trying Hy3 and it kind of drove the point home. This is a model that activates 21B parameters per token out of a 295B total. It's tiny compared to DeepSeek's 671B or Kimi K3's 2.8 trillion. And yet for the coding and API integration work I threw at it, the output quality was closer to those models than it had any right to be. That's the part that's hard to ignore.

When DeepSeek alone is dominating OpenRouter usage, Qwen is leading Arena-Hard, and now even a small efficiency-focused model like Hy3 is hanging with them on real tasks…the "moat" around OpenAI and Anthropic just doesn't match what I'm seeing day to day. If this is what a 21B-active model can do in mid-2026, I genuinely don't know what the gap argument is even based on anymore.

However that’s just my feelings, I’m curious what everyone else is seeing in their own stacks.