Skip to main content
  1. Home
  2. Computing
  3. News

This AI creativity study says you still beat it, if you’re top tier

A huge new comparison finds some models beat average scores, but humans dominate the high end.

Add as a preferred source on Google
Light, Lightbulb, Can
Omar:. Lopez-Rincon

Generative AI just cleared a new bar in creativity, at least for the average person. This AI creativity study compared results from more than 100,000 people with several large language models, including ChatGPT, Claude, and Gemini, and it found some models can outscore a typical human on a standardized creativity task.

But the ceiling still looks human. The study reports the most creative half of participants outperformed every AI model tested, and the top 10% widened the lead even more.

Recommended Videos

AI is getting better at clearing baseline creative tasks, while exceptional human output keeps a gap that’s hard to erase.

The test behind the claim

The researchers leaned on the Divergent Association Task, a quick prompt that asks for ten words that are as unrelated to each other as possible. Scores rise when those words are more semantically distant, and most people finish in a few minutes.

That simplicity is why the team could run such a large comparison. It also helps explain the headline result, models can be tuned to generate wide-ranging word choices on demand, which maps neatly onto what DAT rewards.

Still, DAT measures one slice of creativity, the ability to produce divergent language. It doesn’t measure taste, emotional impact, or whether an idea is the right one for a specific audience.

Where humans keep an edge

The strongest signal in the findings isn’t a single winner, it’s the spread. Some AI systems can beat the middle of the pack, but high-scoring humans separate themselves, and the separation grows at the top end.

In day-to-day terms, models excel at volume. If you need ten directions fast, it can deliver. What it can’t reliably do is the selective part, choosing the one direction worth pursuing, shaping it for constraints, and making it feel intentional instead of merely plausible.

That’s also why the result shouldn’t be read as a verdict on creative careers. The benchmark shows ideation range. It doesn’t show judgment under pressure, or the kind of originality that changes what an audience expects.

What to do with it

The team also compared people and models on creative writing style tasks, including haiku, plot summaries, and short stories, which better resembles how many people use ChatGPT. Even there, top human creators kept the advantage.

If you’re using AI at work, treat it as an ideation accelerator. Use it to generate breadth, then apply the part that still separates you, decide what fits your voice, what matches the brief, and what’s worth shipping.

Keep an eye on follow-ups that pin down exact model versions and test dates, because this kind of leaderboard can move quickly as models change.

Paulo Vargas
Paulo Vargas is an English major turned reporter turned technical writer, with a career that has always circled back to…
Claude Opus 5 is here, and Anthropic says it can rival Fable 5 in some tasks
Major software engineering improvements put Opus 5 closer to Anthropic’s top model
Claude Opus 5 logo

Anthropic has launched Claude Opus 5, its latest frontier AI model for coding, research, business work, and other complex tasks. The company says it delivers a major performance jump over Claude Opus 4.8 while keeping the same API price. Anthropic also claims it comes close to Claude Fable 5 on some coding and computer-use tests while costing far less per task.

Opus 5 is available across Claude’s apps and API. It is now the default model for Claude Max subscribers and the strongest option included with Claude Pro.

Read more
Humans actually prefer talking to an AI than a support person, says gas giant as it cuts jobs
British Gas cuts 1,300 support jobs as its CEO points to changing customer habits
Executive, Person, Electronics

Centrica, owner of British Gas, is cutting 1,300 call centre jobs, and its chief executive says changing customer behaviour is the main reason. The company plans to remove 800 roles as part of a “targeted deployment of AI tools,” on top of the 500 cuts announced last month. Customer service teams in Glasgow, Edinburgh, Cardiff, Leicester, Stockport, and Leeds will be affected over the next two years.

Some positions will disappear when employees leave and are not replaced, while the remaining cuts will come through redundancies. Trade unions have warned that Centrica’s AI investment will hand hundreds of human jobs to chatbots.

Read more
Intel just pulled its 14A production schedule forward by a year
Risk production for 14A is now planned for late 2027
Intel Core Ultra Desktop CPU

Intel has moved up the schedule for its next-generation 14A (1.4-nanometer) manufacturing process, giving its foundry turnaround an important new target.

During Intel’s Q2 2026 earnings call, CEO Lip-Bu Tan said 14A risk production for internal products is now planned for the second half of 2027. High-volume production is expected to begin in 2028. This is earlier than the previous timeline, when Intel was expected to begin risk production in 2028 and move to volume production in 2029.

Read more