RCreddit.com·
暂不在当前实时榜单
We benchmarked 24 LLMs against human writers on 475 creative writing prompts
A new benchmark has been released, comparing 24 large language models (LLMs) against human writers on 475 creative writing prompts. This benchmark uses a custom reward model trained on human preferences to predict what a large group of readers would prefer. The results indicate that the strongest frontier LLMs already outperform talented amateur writers, though professional writers still maintain a significant lead. The full benchmark and model outputs are available for review.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年9月19日 06:00 UTC
- 收录
- 2026年9月19日 06:00
- 来源类型
- 开发者社区
讨论趋势
暂无对比
百分比基于采集到的讨论信号,不代表新增评论数或独立参与人数。曲线仅用于同一话题在不同时段的比较。
本站未收录正文。
前往源站阅读 →