← Back

Wire · technology

New Benchmark Pits Human Writers Against 24 LLMs Across 475 Prompts, Results Show ...

Published

19 September 2026

Topic

technology

◆ Sectors

Generative AIAI Frontier Models

◆ Geography

United States

◆ Source

Read at wccftech.com →

◆ Verified

Fusion42 · 19 September 2026 · Fusion42 review

A new benchmark by Vulsar AI tested 24 large language models (LLMs) against amateur and professional human writers across 475 creative writing prompts. The results show that only the most advanced frontier models like GPT 6 Astra slightly outperform amateur humans, while professional writers still outperform all AI tested.

This Wire brief sits within Fusion42's coverage of Generative AI and AI Frontier Models.

◆ ◆ The Wire takeaway

The ceiling for AI writing quality is rising but professional writers remain ahead. If you build AI content tools, focus now on niches where AI still fails, such as coherent multi-chapter storytelling, before chasing fully professional-level writing.

◆ Coverage

1 source · 19 Sep 2026

◆ Related on Wire

◆ Topics

Generative AIAI Frontier Modelsllmcreative-writingbenchmarkgpt-6human-vs-ai