株式会社極東書店トップ > 商品一覧 > Bandwidth Extension of Speech Using Perceptual Criteria.
商品詳細
Bandwidth Extension of Speech Using Perceptual Criteria.
・ISBN 978-3-031-00393-6 paper EUR 27.99
¥7,481.- (税込) ※(※)価格はご注文時の参考価格となります。
納品価格につきましては書籍の入荷時点で確定となります。
版元の原価改定、外国為替の変動等により異なる場合がございますので、予めご了承下さい。
お気に入り
★★★
| 著者・編者 | Berisha, Visar / Sandoval, Steven / Liss, Julie, |
|---|---|
| シリーズ | (Synthesis Lectures on Algorithms and Software in Engineering) |
| 出版社 | (Springer International Publishing AG, SZ) |
| 出版年月 | 2013 |
| ページ数 | 71 pp. |
| 言語 | ENG |
| ニュース番号 | <A02-83451> |
解説
Bandwidth extension of speech is used in the International Telecommunication Union G.729.1 standard in which the narrowband bitstream is combined with quantized high-band parameters. Although this system produces high-quality wideband speech, the additional bits used to represent the high band can be further reduced. In addition to the algorithm used in the G.729.1 standard, bandwidth extension methods based on spectrum prediction have also been proposed. Although these algorithms do not require additional bits, they perform poorly when the correlation between the low and the high band is weak. In this book, two wideband speech coding algorithms that rely on bandwidth extension are developed. The algorithms operate as wrappers around existing narrowband compression schemes. More specifically, in these algorithms, the low band is encoded using an existing toll-quality narrowband system, whereas the high band is generated using the proposed extension techniques. The first method relies only on transmitted high-band information to generate the wideband speech. The second algorithm uses a constrained minimum mean square error estimator that combines transmitted high-band envelope information with a predictive scheme driven by narrowband features. Both algorithms make use of novel perceptual models based on loudness that determine optimum quantization strategies for wideband recovery and synthesis. Objective and subjective evaluations reveal that the proposed system performs at a lower average bit rate while improving speech quality when compared to other similar algorithms.