Résumé
A novel time-domain algorithm is presented for text-to-speech synthesis using diphone concatenation. The algorithm is based on the pitch-synchronous overlap-add (PSOLA) approach and is capable of good quality prosodic modifications of natural speech. The algorithm can be seen as a simplification of a previous algorithm combining the PSOLA approach and frequency-domain transformations. On the other hand, it appears as a generalization of previous time-domain methods that perform pitch synchronous cut-and-splice operations on the speech waveform. This algorithm is used in the CNET diphone synthesis multilingual system, actually supporting three languages: French, Italian, and German. The resulting speech has been tested on French and is judged of much better quality than for an LPC-based synthesizer.
| langue originale | Anglais |
|---|---|
| Pages (de - à) | 238-241 |
| Nombre de pages | 4 |
| journal | ICASSP, IEEE International Conference on Acoustics, Speech and Signal Processing - Proceedings |
| Volume | 1 |
| état | Publié - 1 déc. 1989 |
| Modification externe | Oui |
| Evénement | 1989 International Conference on Acoustics, Speech, and Signal Processing - Glasgow, Scotland Durée: 23 mai 1989 → 26 mai 1989 |
Empreinte digitale
Examiner les sujets de recherche de « Diphone synthesis system based on time-domain prosodic modifications of speech ». Ensemble, ils forment une empreinte digitale unique.Contient cette citation
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver