Daniel Miessler • 8/9/2025

The Worst AI Metric

The article argues that the 'how many r's in strawberry' test is a terrible metric for evaluating AI. It compares it to interrupting a human writer mid-sentence to ask about vowel counts, which disrupts the creative process. The author contends we should judge AIs (and humans) on the quality of content they produce, not on simultaneous meta-analysis of that content, noting that AI tokenization makes the test even less relevant.

0 comments

#artificial intelligence #ai #metrics