No thanks.

I'm good with DeepSeek v4.1 set to high. It is a relentlessly "hardworking" dirt cheap model.

Told it to convert a products page (that had two different fonts based on language) from two columns layout to 5 columns on desktop and 2 columns on mobile ensuring typography is readable.

My man went into spawning sub agent which failed to drive chrome so it wrote its own chrome driver protocol server in Typescript then generated a prototype website then downloaded the images and rendered each variation in a directory taking 100+ screenshots analyzing the typography depth and then delivering detailed report and then writing the whole thing with new page layout testing it again with several dozen screenshots using its driver and then saying all good and all really was good and whole thing took 25 minutes or so (including double visual validation) because it generates token at an incredible speed.

Total cost of the above? $0.07 cents.

PS: It generates token at such a blazing fast speed that you can't recognize the words as they are being added and can't read it without scrolling and pausing even if you're Jimmy Carter.

Also include that all of this comes with full reasoning traces, so if something goes wrong, you know exactly what assumption it started from.

Yes exactly. Reading this "thinking" traces is a great tool.

Assuming you are not using DeepSeek with Claude code so what/how are you using it?

Same here! DeepSeek v4.1 Flash has been my moment of "does everything I need, cheaply. Please now focus all R+D on making this efficient enough to run off a laptop"

Competition is good

It really is good. I forgot to mention that within that said sub agent, it also went into exploring top e-commerce websites (Zalaondo, Temu, Amazon, eBay) for exploring prevailing industry UX best practices and taking screenshots of their product and category pages with its own written chrome driver that I talked about and then went onto prototyping a new website in a temporary directory and then taking hundreds of screenshots to analyse what would be the best column density one each medium for each language.

And that all is 0.07 cents all included.

What harness do you use with it? Are you using v4.1flash via open router ?

I am using DeepSeek Harness[0] (switched from OpenCode) and I am using DeepSeek directly via the API. The speed is insane. Like 200 tokens/second is the norm but I have seen much higher too at times.

PS: I do not know why but opencode pushes CPU usage to very high which has NOT happened with DeepSeek harness even once.

[0]. https://github.com/deepseek-ai/deepseek-harness

Compared to Opus 5, and others, I also found DeepSeek 4.1 Max to be really good and cheap. I am testing right now with Opus 5.5 and I feel it way cheaper than Opus 5!