What tools do we have to countermeasure the state sponsored bias in the Chinese models? Doesn’t seem like a smart plan if individuals can just compensate for the bias.

Also, the choice right now is between an open weight Chinese model that is hypothetically censored to block / sabotage routine engineering flows vs a closed weight service that is definitely censored to block / sabotage those things.

First anthropic guardrails blocked totally normal stuff on fable and knocked you down to opus. At this point, they kick you off fable, then opus, then sonnet. Claude then automatically builds up memories of techniques to bypass the guardrails in my long running sessions (the coordinator agent notices the subordinates got shot in the head and their sessions were pulled from context, so it parses out the lost context from ~/.claude json files, then reformulates parts of the task and uses partial results until the guardrail doesn’t trip.

If I were paying for the API, this dance would cost $50-100 a pop, but I’m not, so whatever (for now).

I worry that AI will be so fundamental to how we do things in the future, companies can mold human behavior via access to the AI tools.

For instance, I worked at FICO. When I mention this, people wonder what they do. The average person only knows FICO as a “score.”

FICO was founded in Silicon Valley.

The average person doesn’t think about how credit scores work, fundamentally.

It’s just software, at its core. FICO incentivizes certain behaviors.

Ever been banned from an online forum?

Now imagine if a corporation could shut you off from a technology that’s literally indispensable.

Same idea.