Skip to content
RCreddit.com·
Not on the current live radar

We benchmarked 24 LLMs against human writers on 475 creative writing prompts

AI summary

A new benchmark has been released, comparing 24 large language models (LLMs) against human writers on 475 creative writing prompts. This benchmark uses a custom reward model trained on human preferences to predict what a large group of readers would prefer. The results indicate that the strongest frontier LLMs already outperform talented amateur writers, though professional writers still maintain a significant lead. The full benchmark and model outputs are available for review.

Time & source

Times shown in UTC

Display time zone: UTC

Local time zone unavailable; showing UTC.

IngestedOffset at this time: UTC+0Sep 19, 2026, 06:00 UTC

Ingested
Sep 19, 2026, 06:00
Source type
Dev community

Discussion trend

No comparison yet
Latest 24h versus previous 24h snapshot means · 7-day curve

The percentage is based on collected discussion signal, not new comments or independent people. The curve only compares the same topic across time.

Full text isn't available here.

Read at source →
Source·reddit.com