How I AI · Tuesday, June 30, 2026
A key component of the 'How I AI' benchmark is the evaluation of an agent's voice and personality, with the host noting a preference for Sonnet 46's conversational style. This subjective scoring aims to assess the 'vibe' of the interaction, beyond mere task completion.
“I am very picky about the personality of my agents. And in particular, the personality of my open clock.”
“And so one of my checks was: given a model, how is its voice? Do I want to hang with it?”