Strawberries are a well known weakness of LLMs, as they have a hard time to count the numbers of "r"s in them. Maybe that's why, because they fear that weakness could be exploited somehow.

Strawberries aren't the weakness, the weakness is the tokenization of a prompt. Any word with multiple duplicate characters is going to be troublesome for LLMs.

Probably need to be taught by someone of Latino origin. Learning to roll them "r"s could help.

2/3rds of all languages use rolled r's

2/3rrrrds of all languages use rolled r's

Really? I find that hard to believe and can't find a reference for it.

I always thought this should be easier by telling it to write each letter one a new line, then count, any token separator ought to suffice, so much so you'd think they'd have trained in this strategy given tokens make individual letters opaque