According to information published on the Hugging Face blog, existing benchmarks indicate that voice artificial intelligence is approaching the level of human performance. However, real-world conversations tell a different story. This may mean that current evaluations do not fully reflect the realities of using voice models in everyday life.

The new Real World VoiceEQ metric, introduced by Hugging Face, aims to more accurately assess the quality of voice artificial intelligence from the perspective of human perception. This could help identify shortcomings not accounted for in standard benchmarks and improve the application of voice models in real-world conditions.