I've been driving flash model for 90% of my tasks. It's better than pro (for unknown reasons), very cheap and fast.

I try to keep changes under 1000 lines and drive architectural decisions myself, barely notice any difference compared to frontier models. The rest 10% is to spot bugs, security problems and to investigate better architecture, which flash can also do pretty well, I just cross check it.

Faster iterations are way better for me, I hate waiting for 5-10 minutes on small changes. I tried to use recent versions of Kimi and GLM, but they use too much thinking for no reason and are pretty slow because of it. I also often feed a lot of data to it, without worrying about hitting the limits: dependencies (to find bottlenecks in them), logs, performance dumps and so on.

Also, it will never complain about security guards, I've been using it to reverse engineer binaries.

> Also, it will never complain about security guards, I've been using it to reverse engineer binaries.

Maybe I'm using too weak language in my prompts, but none of the OpenAI models I've used via codex has refused to reverse engineer binaries, is it supposed to? I'm sitting right now reverse-engineering a 3rd party firmware together with Codex and haven't hit a single guardrail. Meanwhile, I see people complaining about it rejecting non-security related prompts, are things so individual on the platforms right now or what's going on?

Have you completed the identity verification? It's much more lenient once you have

Oh yeah, back in the GPT3.5 days I think, that's probably it. Thanks for sharing your hunch :)

Weird, I'm using 5.6 Sol through Chinese resellers and it reverse engineers stuff just fine

Because they have gone through identify verification, they are more likely to have done that to increase reputation score.

That's hilarious that Anthropic and OpenAI can't even secure their products from Chinese resellers, how are they supposed to secure their models from being distilled?

They can't, that's the entire market for resellers all the data collected is used for distillation

I think I’m missing something, what do you mean Chinese resellers? As I understand it, it’s difficult to even access OpenAI in China, how could they be reselling it? Do you mean something like openrouter but a Chinese version or something?

They usually resell codex subscriptions as api so it's cheaper than the official api

Proxies/vpns to enter/exit comms through non-banned countries.

Do you have any suggestions for resellers? My Gmail username is the same as my HN username if you're not comfortable posting that here.

Thanks.

Probably not real 5.6 Sol

How so? The code and reasoning patterns certainly match, so does the intelligence level. I probably wouldn't know if they're secretly serving Luna instead of Sol, but it's definitely an OpenAI model.

I got an account warning on OpenAI (waved after I complained) just because I was asking it how to root some >10 years old Android device.

Sonnet recently even suggested I root a three year old device and provided instructions. I suppose that is depends on use case. My use case was exporting data from an abandoned application running on an S24 Ultra.

The idea was that I could continue to use the application in the future and export the new data, not that I would be able to recover the extent data already in there.

This is my experience also: DeepSeek v4 flash is good enough for most of my work and I like the fast response times. I buy tokens mostly from FireWorks.ai in the US, but I also prepaid for a large chunk of tokens directoy with DeepSeek.

I use OpenCode mostly (uses fewer tokens than Claude Code) and I am looking forward to the release of DeepSeek’s own coding harness.

It's good but (at least on openrouter) it's got an annoyingly tight output token limit. So if it does get stuck in a reasoning pit, it won't work its way out of it in time.

It's replaced the Kimi models for me though.

Try to use DS platform directly - cheaper and better than openrouter, no subscription

Openrouter has cheaper inference than deepseek for the prior version of v4 flash, which I suspect will happen within a day or two with this version, and they don’t offer subscriptions. Are you confusing it with opencode?

No, maybe I've checked too much time ago, but direct DS was cheaper for all versions - also I had issues with openrouter providers availability

There are numerous providers offering full weight versions for a significant discount.

Availability and rate limiting can be an issue, though. I’ve found that constraining the providers works well to solve that though.