Wire · technology
New Benchmark Pits Human Writers Against 24 LLMs Across 475 Prompts, Results Show ...
◆ Sectors
◆ Geography
◆ Source
◆ Verified
Fusion42 · 19 September 2026 · Fusion42 review
A new benchmark by Vulsar AI tested 24 large language models (LLMs) against amateur and professional human writers across 475 creative writing prompts. The results show that only the most advanced frontier models like GPT 6 Astra slightly outperform amateur humans, while professional writers still outperform all AI tested.
This Wire brief sits within Fusion42's coverage of Generative AI and AI Frontier Models.
◆ ◆ The Wire takeaway
The ceiling for AI writing quality is rising but professional writers remain ahead. If you build AI content tools, focus now on niches where AI still fails, such as coherent multi-chapter storytelling, before chasing fully professional-level writing.
◆ Coverage
1 source · 19 Sep 2026
◆ Related on Wire
◆ Topics