← Front page

How I AI · Tuesday, June 30, 2026

Agent Personality and Voice Evaluated in Benchmark

A key component of the 'How I AI' benchmark is the evaluation of an agent's voice and personality, with the host noting a preference for Sonnet 46's conversational style. This subjective scoring aims to assess the 'vibe' of the interaction, beyond mere task completion.

The tape

2 quotes
I am very picky about the personality of my agents. And in particular, the personality of my open clock.
And so one of my checks was: given a model, how is its voice? Do I want to hang with it?
Heard on How I AI — “Sonnet 5 review: I ran 64 generations to find out if it's worth it, published Tuesday, June 30, 2026. Heardvine summarizes and quotes with attribution and timestamps, and links to the original everywhere.
Transcribed via Gemini audio transcription · $0.02
Agent Personality and Voice Evaluated in Benchmark — Heardvine