Hume unveils Octave, a groundbreaking AI model that goes beyond reading text: it grasps its meaning, producing natural, expressive voices that capture emotions and contexts like never before.
Hume has introduced Octave, a text-to-speech system bringing a fresh approach to artificial intelligence. Unlike conventional methods that simply pronounce words, this model—described by its creators as the first large language model for text-to-speech—interprets a text’s context and emotions. It adjusts tone, rhythm, and timbre, delivering whispers for intimate scenes or calm explanations, much like an actor reading a script.
In a test with 180 evaluators, Octave outperformed ElevenLabs, a notable competitor. It earned 71.6% preference in audio quality, 51.7% in naturalness, and 57.7% in matching voice descriptions, based on 120 varied examples, from movie narrators to medieval characters. These results highlight its ability to adapt to diverse styles and needs.
The system features tools like Voice Design, which crafts unique voices from detailed descriptions, such as an empathetic counselor or a medieval knight. It also offers Acting Instructions, enabling real-time tweaks to emotions and styles. Soon, it will add voice cloning, requiring just five seconds of audio to replicate a voice.
Octave is now accessible on platform.hume.ai and via API, making it suitable for audiobooks, podcasts, or interactive apps. Alongside this, Hume has launched Expressive TTS Arena, a public platform where anyone can compare advanced voice systems and test their skills with complex, expressive texts.
Developed initially for English and Spanish, Octave is still evolving. Beyond synthesizing speech, it explores how people express themselves, paving the way for future AI applications.
Research laboratory and technology company specialized in AI models with emotional intelligence. Its main model integrates voice and language processing, with adjustable voice synthesis in timbre, ...
12/06/2026
The United States government has ordered Anthropic to block access to Claude Fable 5 and Mythos 5 for foreign nationals, forcing the company to ...
09/06/2026
Anthropic introduces Claude Fable 5 and Claude Mythos 5, two versions of its most capable model to date. They share the same foundation, but one is ...
02/06/2026
Microsoft expands its artificial intelligence portfolio with seven models developed entirely by its MAI team, covering image generation, ...
25/05/2026
Pope Leo XIV publishes the first encyclical dedicated to artificial intelligence, setting human dignity as the criterion for all technological ...