Atlas

Benchmarks

← All benchmarks

LLM Divergent Thinking Score

Chat & Writing · 2025-03-20

Lech Mazur's Divergent Thinking Creativity Benchmark asks models to generate many unique words under letter and association constraints. The headline score rewards semantic distance and valid diversity.

Top models (higher is better)

ModelScore
o1 Preview4.8
Gemini 2.0 Flash Experimental4.7
Opus 34.5
Grok 24.5
Llama 3.3 70B Instruct4.4
Gemini 2.0 Flash Thinking Experimental 01-214.4
Claude 3.5 Sonnet (Oct 2024)4.4
Gemma 2 27B4.4
o1-mini4.2
Haiku 3.54.2
Mistral Large 2 (Instruct 2407)4.1
GPT-4o Mini4.1
Gemini 1.5 Flash4.1
Gemini 1.5 Pro 0024.1
Claude 3 Haiku4.0
Qwen2.5 72B Instruct3.9
Llama 3.1 405B3.8
DeepSeek-V2.53.8
GPT-4o3.7
Loading Atlas data…