It is disclosed an Intonation-Aware Subsonic
Harmonics (ISH) unit (2), adapted to be integrated into a Text-to-Speech (TTS)
system (1). Said unit being comprised by an Intonation Learning Module (2.2), a Subsonic
Harmonics Module (2.3), and a Contextual Feedback Module (2.3). More precisely, the Intonation Learning Module is able to learn the relationship between subsonic
harmonic profiles and intonation patterns across various emotional and contextual scenarios. The Subsonic
Harmonics Module generates and integrates inaudible subsonic frequencies with a synthesized speech waveform (g.), dynamically adjusting them based on the Intonation Learning Module's output and contextual feedback (c.) on
contextual information derivable from the text inputted to the TTS
system, given by the Contextual Feedback Module. This procedure influences prosodic features results in more natural-sounding, fluid, and expressive synthetic speech
signal (h.).