DeepSeek V4 Pro 0813 (Fully Tested): Okay, this is ACTUALLY CRAZY!

AICodeKing10m 36s
0 comments · 0 votesOpen discussionClose discussion
Sign in to join the discussion

    Video summary

    AICodeKing reviews the API release of DeepSeek V4 Pro 0813 and distinguishes official or circulating claims from confirmed details. The video reports strong agentic benchmark scores and unusually low token prices, but cautions that rumored parameter counts and context length were not clearly documented by DeepSeek at the time of recording.

    AICodeKing evaluates DeepSeek V4 Pro 0813 on eight Kingbench tasks: an elevator simulation, interactive contact-lens case, folding-table model, SVG illustration, bow-and-arrow game, permutation problem, autonomous local fine-tuning workflow and dual-time-zone wristwatch. The scores total 61 out of 80, or 76.25 percent.

    AICodeKing reports strong mathematical reasoning, long-horizon execution, planning, front-end work and three-dimensional interfaces, but also identifies meaningful weaknesses. On simple tasks, the model can overthink and reach a worse answer than a smaller variant, while straightforward fixes can trigger unrequested abstractions and broad edits.

    Original YouTube thumbnailWatch on YouTube