← Back to KHAO

DeepSeek · Claude · GPT · Gemini ·

DeepSeek's New Model Nearly Matches GPT-6 Astra on Design—at 1.4% of the Cost

2 min read

Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.

★ Tier-1 Source

Image accompanies the article at Decrypt. No description was extracted from the source.

OpenDesign, the company behind the benchmark site OpenDesign Arena, ran 13 AI models through the same batch of design tasks this week.

Key facts

Summary

OpenDesign Arena scored DeepSeek V4.1 Flash at 81.2 out of 100 on real-world design tasks, 98% of GPT-6 Astra's 82.7, while charging $0.023 per finished design against Astra's $1.61. Of the 13 models tested—including Claude Fable 5.1, Grok 4.6, and Qwen 3.8-Max—11 scored lower than DeepSeek's model and cost more to run. DeepSeek's technical paper for V4.1 Flash shows the model activates 8 billion of its 552 billion parameters to read a prompt, the design choice behind its low price. OpenDesign Arena scores models on everyday design work—building web apps, dashboards, mobile screens, and landing pages—out of 100 points.

Read full article at Decrypt →

#DeepSeek #Claude #GPT #Gemini