LATENT REFERENCES / TAG2
Phase Vocoder
Original title: フェーズボコーダ
This reference note belongs to Tag2 in Latent References, an archive curated by Keigo Yoshida. Its archive region is Phase coherence. The note preserves its source text and links so that readers can trace the material behind the 3D map.
- Collection
- Tag2
- Archive region
- Phase coherence
Archived reference note
English translation of the archived note. JP shows the original text. Source links and literal code are retained; the translation does not update or independently verify the source claims.
(English: Phase vocoder) is a vocoder modeling audio signals through amplitude and phase in the frequency domain.
At the heart of the phase vocoder is the short-time Fourier transform (STFT), through the following stages.
Analysis: Conversion from time-domain representation → time–frequency representation (English version) by STFT.
Modification: Manipulating amplitude and phase of arbitrary frequency components.
Resynthesis: Frequency-domain representation → time-domain representation through inverse STFT.
A phase vocoder allows time-stretching and pitch conversion of audio signals through frequency-domain modification. Changing STFT analysis frames' temporal positions before resynthesis also changes the resynthesized result's evolution in time, for example enabling changes in sound timescale.
British composer Trevor Wishart (English version) created “Vox V()” (on the album “Vox Cycle (English version)”) based on phase-vocoder analysis/transformation of the human voice. American composer Roger Reynolds's “Transfigured Wind()” used a phase vocoder to time-stretch flute sounds.
Introduced as an algorithm maintaining horizontal coherence between phases of bins representing sinusoidal components. This original phase vocoder did not consider vertical coherence between adjacent frequency bins, so audio signals time-stretched by the system lacked clarity.
(英語: Phase vocoder)は音声信号を周波数領域の振幅と位相でモデル化するボコーダである。
フェーズボコーダの心臓部は短時間フーリエ変換 (STFT)であり、次の段階を経る。
分析: STFTによる時間領域表現→時間-周波数表現(英語版)変換
変更: 任意の周波数成分の振幅・位相操作
再合成: 逆STFTによる周波数領域表現→時間領域表現変換
フェーズボコーダは周波数領域での変更処理により音声信号の時間伸縮とピッチ変換などを可能にする。また再合成前にSTFT分析フレームの時間的位置を変更すれば、再合成結果の時間発展を変更でき、たとえば音の時間スケール変更を実現できる。
イギリスの作曲家 トレヴァー・ウィシャート(英語版)は、人間の声のフェーズボコーダ分析/変換に基づいて、“Vox V()” (アルバム “Vox Cycle(英語版)”) を制作した。アメリカの作曲家 ロジャー・レイノルズの作品 “Transfigured Wind()” は、フェーズボコーダをフルート音のタイムストレッチに使用した
正弦波成分を表す各ビンの位相間で水平コヒーレンスを維持するアルゴリズムとして導入された。このオリジナルのフェーズボコーダは、隣接する周波数ビン間の垂直コヒーレンスを考慮しなかったので、このシステムによるタイムストレッチ(時間伸縮)の音響信号は明瞭さが欠けていた。
Source updated 2023-06-15 · Snapshot 2026-10-08
Source links and calculated neighbors
Cosine values measure shared lexical features, not truth, agreement or identical meaning. Original reference links are labeled separately.
- Phase Coherence ProblemComputed lexical cosine similarity 0.257 · shared title, text, tags and references
- Phased ArrayComputed lexical cosine similarity 0.129 · shared title, text, tags and references
- Plume CoherenceComputed lexical cosine similarity 0.102 · shared title, text, tags and references
- CauseComputed lexical cosine similarity 0.090 · shared title, text, tags and references