AI Disruption

AI Disruption

Claude Code: Best Harness for DeepSeek-V4.1-Flash

Claude Code and DSH Standard lead DeepSeek-V4.1-Flash harness tests, beating Minimal and Pi Agent in speed, tokens, and task quality.

Meng Li's avatar
Meng Li
Sep 15, 2026
∙ Paid

“AI Disruption” Publication 10,500 Subscriptions 20% Discount Offer Link.


I Tried the (New) Claude Code Rival DeepSeek Harness (That's Wild & FREE) |  by Joe Njenga | Aug, 2026 | Medium

In my analysis of DeepSeek-V4.1-Flash, I noticed that DeepSeek’s technical documents mention this: large models rarely run inside only one fixed harness. Different scaffolds differ in system prompts, tool definitions, context management, and interaction protocols, and the results they produce also differ.

DeepSeek V4.1 Flash Released, Now Adapted for Harness

DeepSeek V4.1 Flash Released, Now Adapted for Harness

Meng Li
·
Sep 10
Read full story
Image

DeepSeek’s official tests show:

  1. In DeepSeek’s own framework, DSH Minimal mode scored higher than Standard and PTC on both tests. That is a bit counterintuitive: more features and a more complex workflow did not yield a higher success rate.

  2. When DeepSeek-V4.1-Flash is plugged into Claude Code or Codex, the out-of-the-box level is already close to peak performance.

I ran a practical test:

Write an HTML and JavaScript page implementing Space Invaders

I ran Pi Agent, Claude Code, DSH Standard mode, and DSH Minimal mode separately. First, look at the results; time, token usage, cache hit rate, and output speed will be analyzed later.

User's avatar

Continue reading this post for free, courtesy of Meng Li.

Or purchase a paid subscription.
© 2026 Meng Li · Privacy ∙ Terms ∙ Collection notice
Start your SubstackGet the app
Substack is the home for great culture