← all stories

Deepseek V4 Pro

2h ago

DeepSeek V4 Flash Stumbles on Real Agent Tasks as Prices Surge

DeepSeek's V4 Flash model, popular for its low cost and high benchmark scores, is facing criticism after third-party tests showed it struggling on real-world agent tasks. Composio ran 30 difficult multi-step workflows across four harnesses, with only 6 fully succeeding. The company also raised API prices on August 6, 2026, citing unprecedented demand.

4d ago

DeepSeek Quietly Releases V4 Pro 0813 Model

Chinese AI startup DeepSeek has released its latest AI model, DeepSeek V4 Pro 0813, without a formal announcement. The release was first noticed on social media and message boards, where users and tech commentators highlighted the model's performance and cost efficiency. According to the open-source AI coding agent Cline, the model scores 15.8 percent higher on Terminal Bench, a benchmark that evaluates how accurately AI can execute real system operations and development tasks on a Linux terminal, compared to the April preview model. Cline also noted that DeepSeek V4 Pro 0813 runs at roughly one fifty-seventh the cost of Claude Fable 5 while delivering comparable performance. The model has 1.6 trillion parameters, 49 billion active parameters, and a 1 million token context window. Tech writer @ChrisGPT reported benchmark improvements across Terminal Bench 2.1, CyberGym, and DeepSWE, with scores rising from 72.1 to 87.9 percent, 52.7 to 83.3 percent, and 12.8 to 62.7 percent respectively. Pricing is set at 0.435 dollars per million input tokens and 0.87 dollars per million output tokens, with a discounted cache-hit input rate of 0.003625 dollars.