"Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok
Researchers tested four frontier AI models by having them draw famous artworks like the Mona Lisa using a shared toolset. The study highlights how different models interpret open-ended creative tasks and provides a framework for evaluating their visual output.
Why it matters
This experiment helps users understand the varying creative capabilities and tool-use behaviors of current generative AI models beyond standard text-based benchmarks.
All posts comparison agents drawing "Drawing" the Mona Lisa with GPT-5.6, Claude, Gemini, and Grok Four frontier models, a blank canvas, and colored pencils. We tracked every stroke, dollar, and output as they tried to draw the Mona Lisa.
We ran four vision models, GPT-5.6 Sol , Claude Fable 5 , Grok 4.5 , and Gemini 3.6 Flash , across two targets (the Mona Lisa and Van Gogh's Starry Night, both scored objectively) and five open-ended prompts, for 28 drawings total. We'll cover tool use, cost, output, and whether the models actually improved their work, with our opinion at the end.
Get smarter about the news
Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.
Create free accountAlready have an account? Sign in