Spectral modeling synthesis

Spectral modeling synthesis (SMS; スペクトラルモデリング合成) は正弦波と色付きノイズを用いた楽音分析合成手法[2]および音声分析合成手法である[3]

スペクトラルモデリング合成
Roads 1996, p. 153[1]を日本語訳

概要

SMS分析/合成の処理概要
    Bonada et al. 2001 Fig.1 & 2に基く

SMSは調波成分残余成分 (非調波成分; ノイズ成分) の組合せとしてモデル化する。

楽音分析合成 / 音声分析合成として次の要素から構成される。

このモデルは多くのタイプのオーディオ信号に適用できる。例えば音声信号は、声帯振動で生じるゆっくり変化する調波音と、唇や口で生じる広帯域ノイズ状音を含む。同様に楽器も、調波成分と、ノートの発音/変更時に生じるノイズ状音の両方を発する。

関連項目

Sinusoidal modeling
Sinusoidal Analysis/Synthesis System
(McAulay & Quatieri 1988, p. 161[5] に基く)

脚注

  1. Roads 1996, p. 153, Figure 4.23: Overview of spectrum modeling synthesis.
  2. 本手法は調波解析/調波合成に基づいており、その意図は調波成分が主役となる楽音音響分析音響合成である。
  3. Serra & Smith 1990, p. 12. "It describes a technique called spectral modeling synthesis [SMS], that models time-varying spectra as (1) a collection of sinusoids controlled through time by piecewise linear amplitude and frequency envelopes (the deterministic part), and (2) a time-varying filtered noise component (the stochastic part). The analysis procedure first extracts the sinusoidal trajectories by tracking peaks in a sequence of short-time Fourier transforms. These peaks are then removed by spectral subtraction. The remaining “noise floor” is then modeled as white noise through a time-varying filter. A piecewise linear approximation to the upper spectral envelope of the noise is computed each successive spectrum, and the stochastic part is synthesized by mean of the overlap-add technique."
  4. 加法性ホワイトガウスノイズ (AWGN): パワースペクトル(周波数領域の強度)が全周波数で同じ強度(=白色)で、振幅分布(時間領域の強度)がガウス分布に従うノイズ
  5. McAulay & Quatieri 1988, p. 161, Fig. 8. "This block diagram of the sinusoidal analysis/synthesis system illustrates the major functions subsumed within the system. Neither voicing decisions nor residual waveforms are required for speech synthesis."

参考文献

外部リンク

This article is issued from Wikipedia. The text is licensed under Creative Commons - Attribution - Sharealike. Additional terms may apply for the media files.