01
🎙️ Voice Human-Likeness
You can now score how human vs robotic your agent actually sounds. Voice Human-Likeness is a new 1-5 metric that listens to the agent's own audio (not the transcript) and judges acoustic delivery against a fixed rubric.

What it gives you:
- Per-agent scores sampled evenly across the call, so scripted monologues never dominate the read.
- Catches smooth robots: TTS voices that pass classic MOS checks but still sound robotic to a real listener.
- Pairs with Voice Naturalness, also new this week. Naturalness (UTMOS-based) grades audio signal quality; Human-Likeness grades whether the delivery reads as human. Together they separate "sounds bad" from "sounds fake."


