Edge Rewrite
// HTMLRewriter · presentation

This page was redesigned at the edge.

Cloudflare fetched the original article and streamed it through HTMLRewriter to apply an entirely new visual system without rebuilding the source page.

Jump to content

Fujisaki model

From Wikipedia, the free encyclopedia
An F0 contour is obtained by adding the phrase and accent components to the base frequency

The Fujisaki model is a superpositional model for representing F0 contour of speech.

As a model of intonation, it was a major milestone in prosody modeling for speech synthesis.[1]

According to the model, F0 contour is generated as a result of the superposition of the outputs of two second order linear filters with a base frequency value. The second order linear filters are for generating the phrase and accent components of speech. The base frequency is the minimum frequency value of the speaker. In other words, F0 contour is obtained by adding base frequency, phrase components and accent components. The model was proposed by Hiroya Fujisaki.


where

Where,

: bias level upon which all the phrase and accent components are superposed to form an contour,

 : number of phrase commands,

 : number of accent commands,

 : magnitude of the ith phrase command,

 : amplitude of the jth accent command,

 : instant of occurrence of the ith phrase command,

 : onset of the jth accent command,

 : end of the jth accent command,

 : natural angular frequency of the phrase control mechanism to the ith phrase command,

 : natural angular frequency of the accent control mechanism to the jth accent command, and

 : ceiling level of the accent component for the jth accent command.

References

[edit]
  1. ↑ Gibbon, Dafydd (2017-04-27). "Prosody: The Rhythms and Melodies of Speech". arXiv:1704.02565 [cs.CL].

Bibliography

[edit]