Interestana
Home/News/DeepSeek V4 Pro Benchmarks Show Marginal Gains Over Claude Fable
Decrypt2 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

DeepSeek V4 Pro Benchmarks Show Marginal Gains Over Claude Fable

DeepSeek V4 Pro Benchmarks Show Marginal Gains Over Claude Fable

Chinese AI company DeepSeek has released its V4 Pro large language model, which benchmarks indicate offers only marginal improvements over Anthropic's Claude Fable. In an April preview, DeepSeek V4 Pro scored 18 points lower than Claude Fable on a key benchmark, suggesting a significant performance gap. However, DeepSeek's own internal benchmarks for the finished V4 Pro model present a contrasting narrative, claiming superior performance. This discrepancy raises questions about the validity and transparency of the benchmark results presented by DeepSeek.

Further analysis of the V4 Pro's performance reveals a substantial price increase. While the exact cost of the V4 Pro has not been fully disclosed, reports indicate a staggering 4,500% price hike compared to previous versions or competitor models. This aggressive pricing strategy, coupled with performance metrics that do not demonstrably surpass existing leading models like Claude Fable, positions V4 Pro as a potentially less attractive offering for many users. The company's previous model, V3, was noted for its competitive pricing and strong performance, making the V4 Pro's new cost structure a significant departure.

DeepSeek, headquartered in China, has been actively developing large language models with a focus on open-source contributions and commercial applications. The company aims to compete with global AI leaders by offering powerful models that can be integrated into various business solutions. The V4 Pro is intended to enhance capabilities in areas such as code generation, complex reasoning, and multilingual understanding. The company's strategy often involves releasing models with different parameter sizes and performance tiers to cater to a diverse market, from researchers to enterprise clients.

Anthropic, the developer of Claude Fable, is a prominent AI safety and research company based in the United States. Known for its focus on developing reliable, interpretable, and steerable AI systems, Anthropic has consistently released models that are competitive in performance and often prioritize safety features. Claude Fable, their flagship model, has been recognized for its strong reasoning abilities and its capacity to handle complex tasks, setting a high bar for competing models in the market. The comparison between DeepSeek V4 Pro and Claude Fable highlights the ongoing race among AI developers to achieve state-of-the-art performance while balancing cost and accessibility.

Original source — read the full reporting at the publisher:

Read on Decrypt

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next