<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE ep-patent-document PUBLIC "-//EPO//EP PATENT DOCUMENT 1.1//EN" "ep-patent-document-v1-1.dtd">
<ep-patent-document id="EP94103204B1" file="EP94103204NWB1.xml" lang="en" country="EP" doc-number="0614075" kind="B1" date-publ="20000621" status="n" dtd-version="ep-patent-document-v1-1">
<SDOBI lang="en"><B000><eptags><B001EP>......DE....FRGB..IT..............................</B001EP><B005EP>R</B005EP><B007EP>DIM360   - Ver 2.9 (30 Jun 1998)
 2100000/1 2100000/2</B007EP></eptags></B000><B100><B110>0614075</B110><B120><B121>EUROPEAN PATENT SPECIFICATION</B121></B120><B130>B1</B130><B140><date>20000621</date></B140><B190>EP</B190></B100><B200><B210>94103204.7</B210><B220><date>19940303</date></B220><B240><B241><date>19960111</date></B241><B242><date>19980803</date></B242></B240><B250>en</B250><B251EP>en</B251EP><B260>en</B260></B200><B300><B310>MI930406</B310><B320><date>19930303</date></B320><B330><ctry>IT</ctry></B330></B300><B400><B405><date>20000621</date><bnum>200025</bnum></B405><B430><date>19940907</date><bnum>199436</bnum></B430><B450><date>20000621</date><bnum>200025</bnum></B450><B451EP><date>19990705</date></B451EP></B400><B500><B510><B516>7</B516><B511> 7G 10L  19/04   A</B511></B510><B540><B541>de</B541><B542>Verfahren und Vorrichtung zur Sprachkodierung mit Trellis-kodierter Quantisierung für LPC- Quantisierung</B542><B541>en</B541><B542>Method and apparatus for speech coding using Trellis Coded Quantization for Linear Predictive Coding quantization</B542><B541>fr</B541><B542>Méthode et appareil de codage de parole utilisant la quantisation codée Trellis pour la quantisation LPC</B542></B540><B560><B561><text>EP-A- 0 230 001</text></B561><B561><text>US-A- 4 975 956</text></B561><B562><text>IEEE TRANSACTIONS ON COMMUNICATIONS, vol.38, no.1, January 1990 pages 82 - 93 M.W.MARCELLIN ET AL. 'Trellis Coded Quantization of Memoryless and Gauss-Markov sources'</text></B562><B562><text>ICASSP 91, vol.1, 14 May 1991, TORONTO pages 661 - 664 K.K. PALIWAL ET AL. 'Efficient vector quantization of LPC parameters at 24 bits/frame'</text></B562><B562><text>ADVANCES IN SPEECH CODING, 1 January 1991 pages 47 - 56 M.W. MARCELLIN ET AL. 'A Trellis-searched 16 kbit/sec speech coder with low-delay'</text></B562><B562><text>K. SAM SHANMUGAM: "Digital and Analog Communication Systems" , , JOHN WILEY &amp; SONS, NEW YORK</text></B562><B566EP><date>19960409</date></B566EP></B560><B590><B598>1</B598></B590></B500><B700><B720><B721><snm>Fratti, Marco</snm><adr><str>Via della Birona 9</str><city>I-20052 Monza (Milano)</city><ctry>IT</ctry></adr></B721><B721><snm>Cucchi, Silvio</snm><adr><str>Via S. Ibenzio 9</str><city>I-20090 Gaggiano (Milano)</city><ctry>IT</ctry></adr></B721></B720><B730><B731><snm>ALCATEL</snm><iid>00201873</iid><irf>TLT 342X/ 90135</irf><adr><str>54, rue La Boétie</str><city>75008 Paris</city><ctry>FR</ctry></adr></B731></B730><B740><B741><snm>Knecht, Ulrich Karl, Dipl.-Ing.</snm><sfx>et al</sfx><iid>00070611</iid><adr><str>Alcatel
Intellectual Property Department, Stuttgart
Postfach 30 09 29</str><city>70449 Stuttgart</city><ctry>DE</ctry></adr></B741></B740></B700><B800><B840><ctry>DE</ctry><ctry>FR</ctry><ctry>GB</ctry><ctry>IT</ctry></B840><B880><date>19950802</date><bnum>199531</bnum></B880></B800></SDOBI><!-- EPO <DP n="1"> -->
<description id="desc" lang="en">
<p id="p0001" num="0001">The present invention relates to a method for speech coding as set forth in the preamble of claim 1 and a speech coder as set forth in the preamble of claim 19.</p>
<p id="p0002" num="0002">In the telecommunication field it is useful to transmit the information both vocal and video using as less numbers of bits as possible without loosing part of the information transmitted.</p>
<p id="p0003" num="0003">Such aim is achieved by means of suitable coding techniques. There are a lot of these techniques: each of them has its characteristic features. From the field of the communication theory, and in particular from the modulation theory the Trellis Coded Modulation technique is well known. The Trellis Coded Modulation paradigm, combined with well-known quantization theories, gave rise to the Trellis Coded Quantization (TCQ) algorithm.</p>
<p id="p0004" num="0004">From US patent no. 4, 975, 956 a speech coder is known that employs vector quantization of LPC parameters after conversion from RC to LSP. The gain and pitch are encoded using an adaptive tracking technique with a form of trellis coding. Said form of trellis coding, however, involves a too high computational load, since the generated sequences are compared to find the minimum distortion.</p>
<p id="p0005" num="0005">TCQ is a recent technique for efficient scalar encoding of any source.</p>
<p id="p0006" num="0006">In particular, it has been introduced by M.W. Marcellin, T.G: Fisher, in "Trellis Coded Quantization of memoryless and Gauss-Markov Sources", IEEE Trans. on Communication, vol. 38, No. 1, January 1990, for encoding memoryless Gauss-Markov sources. The feature of the trellis coded quantization approach is the use of a structured codebook with an expanded set of quantization levels. Based on the notion of set partitioning introduced by G: Ungerboek, in "Channel Coding with Multilevel/Phase Signals", IEEE Trans. on Information Theory,<!-- EPO <DP n="2"> --> Vol. IT-28, Jan. 1982, the trellis structure then prunes the expanded number of quantization levels down to the desired encoding rate. The encoder uses the Viterbi algorithm for finding a vector of quantized scalars that is closest (according to a predetermined metric) to the unquantized vector.</p>
<p id="p0007" num="0007">Because of the complexity of the matter, only a good knowledge of both the modulation theory (TCM) and of the quantization theory allows a proper implementation and exploitation of this quantization technique.</p>
<p id="p0008" num="0008">The main object of the present invention is therefore substantially an efficient and effective way how to apply the TCQ technique.</p>
<p id="p0009" num="0009">According to the invention, therefore, the method for speech coding is constructed as set forth in claims 1 or 18 and the speech coder as set forth in claim 19.</p>
<p id="p0010" num="0010">Further features of the invention are explained in the depending claims.</p>
<p id="p0011" num="0011">Embodiments of the invention will now be explained in detail with reference to the accompanying drawings, in which figure 1 is showing a general structure of a Trellis Coded Quantizer; figure 2 is a Trellis Coded Quantizer with variable bit allocation; figure 3 is a TCQ scheme for LSP difference quantization; figure 4 is an updated paths in the TCQ.</p>
<p id="p0012" num="0012">The trellis encoder is completely specified by:
<ul id="ul0001" list-style="dash" compact="compact">
<li>The trellis topology (i.e. the connections among the successive states or, equivalently, the description of the underlying finite-state<!-- EPO <DP n="3"> --> machine).</li>
<li>The quantization levels associated with each state transition.</li>
</ul></p>
<p id="p0013" num="0013">In Figure 1, a possible structure for a TCQ is depicted. The labels assigned to each trellis branch are arbitrary, as well as the quantization level partition.</p>
<p id="p0014" num="0014">This particular structure is depicted for the sake of clearness and may not correspond to a 'physical' one.</p>
<p id="p0015" num="0015">The 4-state trellis is fully connected. The number associated to each state transition (branch) represents the quantization value corresponding to that branch. The trellis is a N-stage one, that is, it is employed for coding a N-component input vector.</p>
<p id="p0016" num="0016">Note that 8 possible scalar quantization values are present; however, each trellis state 'sees' only a 4-value subset, as function of its transition to a future state.</p>
<p id="p0017" num="0017">In principle, an exhaustive procedure should be employed for identifying the best quantized vector with respect to a N-value input vector.</p>
<p id="p0018" num="0018">This means that one should identify each possible trellis path and, according to the quantization value associated to each path step, an overall quantization error should be constructed. In the example depicted in Figure 1, this exhaustive procedure would imply identifying 4<sup>n</sup> quantized vectors and measuring the quantization error for each of them.</p>
<p id="p0019" num="0019">This tremendous amount of computation is avoided by using the well-known Viterbi algorithm as descripted by G.D. Forney, in<!-- EPO <DP n="4"> --> "The Viterbi algorithm", IEEE Trans. on Information Theory, Vol. IT-28, Jan. 1982, that, although operating in a step-by-step fashion, guarantees to find the optimal solution.</p>
<p id="p0020" num="0020">With reference to Figure 1, it is easily understood' that, in case a scalar quantization technique was employed for encoding each input vector component, 3<i>N</i> bits would be necessary.</p>
<p id="p0021" num="0021">On the contrary, by using a TCQ, the quantization process would require 2<i>N</i> + 2 bits (where the binary representation of the TCQ 'winning' best initial -or final- state is taken into account).</p>
<p id="p0022" num="0022">A trellis scheme like the one in Figure 1 is an example of what is generally found in the literature. That is, the topological configuration of the trellis is the same at each quantization step.</p>
<p id="p0023" num="0023">Similarly, the quantization level number is the same at each quantization step.</p>
<p id="p0024" num="0024">This configuration may be not the ideal one in case one needs to quantize a vector whose scalar components have a different 'importance scale', according to a predefined performance criterium. In this case, a different bit/sample number may be necessary for each vector component.</p>
<p id="p0025" num="0025">This problem can be solved in the following way: suppose to start with a given topological configuration for the trellis (that is, suppose to start with the trellis depicted in Figure 1); an increase-decrease in the bit/sample assignment at each quantization step can be obtained as follows.
<ul id="ul0002" list-style="dash" compact="compact">
<li>Bit/sample number increase:<br/>
<!-- EPO <DP n="5"> -->this can be obtained by simply adding one or more parallel transitions to the state branches (that is, we have multiple branches at each state transition). The bit/sample number is thus increased according to the number of parallel transitions that are added. <br/>
Referring to the example in Figure 2, it is immediate to verify that in the first quantization step (Step 1) we have doubled the quantization level number (from 8 to 16) and, accordingly, the bit/sample configuration (i.e. from 2 to 3 bits).
<br/>
The quantization level partition associated to each parallel transition can be derived from the optimal set partition theory, as described in M.W. Marcellin, T.G. Fisher, "Trellis Coded Quantization of Memoryless and Gauss-Markov Sources", IEEE Trans. on Communication, vol. 38, No. 1, January 1990, and G. Ungerboek, "Channel Coding with Multilevel/Phase Signals" , IEEE Trans. on Information Theory, Vol. IT-28, Jan. 1982.
</li>
<li>Bit/sample number decrease:<br/>
This can be obtained by changing the trellis topology in a trivial way. <br/>
In particular, the state transition number can be pruned down to the desired encoding rate.
<br/>
With reference to Figure 2, in the third quantization step (Step 3) the encoding rate is halved (and, obviously, the same is true for the quantization level number) since only a subset of the trellis states can be reached. In the third quantization step, the state transition choice is dichotomic, therefore allowing for a single bit/sample in<!-- EPO <DP n="6"> --> the quantization of the corresponding vector component.
</li>
</ul></p>
<p id="p0026" num="0026">It is worth to note that from the implementation point of view, both the bit/sample increase and the decrease can be easily realized while carrying out the Viterbi algorithm in the encoding process. In particular, the addition of parallel transitions implies an increase in the evaluation of the local transition state metrics. On the contrary, pruning a state transition branch implies the assignment of a corresponding 'infinite' local transition state metric.</p>
<p id="p0027" num="0027">In recent years line spectrum pairs (LSP) representation of LPC parameters has become popular in speech coding applications. The LSP are frequency domain parameters strictly related to the formants: the position of a pair of frequencies gives the position of the formant, while their difference carries information about the width of the spectral peak.</p>
<p id="p0028" num="0028">The ordering property of the LSP parameters can be exploited by quantizing the differences between adjacent LSP frequencies instead of the absolute values of the LSP frequencies.</p>
<p id="p0029" num="0029">A proper bit allocation can be assigned to each LSP difference according to its perceptual importance.</p>
<p id="p0030" num="0030">When applied to the quantization of the LSP differences, the TCQ algorithm proves itself to be particularly effective, since the quantization error accumulated in quantizing the - say - first (<i>i</i> - 1)-th LSP differences can be taken into account in the search of the optimum quantization level for the <i>i</i>-th LSP difference.</p>
<p id="p0031" num="0031">Each trellis state will be assigned a "history path", at each <i>i</i>-th<!-- EPO <DP n="7"> --> trellis stage; each state transition belonging to this history path will correspond to a pointer to the quantization level of the corresponding LSP difference. By adding all the LSP differences of each state history path up to the <i>i</i>-th trellis stage, the <i>i</i>-th quantized LSP can be reconstructed (note that this reconstructed LSP will be different - in general - for each trellis state).</p>
<p id="p0032" num="0032">As an example, suppose that the <i>i</i>-th LSP difference must be quantized.</p>
<p id="p0033" num="0033">The following operations will be performed:
<ul id="ul0003" list-style="dash" compact="compact">
<li>For each <i>j</i>-th state of the(<i>i</i> - 1)-th trellis stage, the corresponding (<i>i</i> - 1)-th LSP is reconstructed, by adding all the quantized LSP differences belonging to the <i>j</i>-th state history path.</li>
<li>For each <i>j</i>-th state, the LSP difference between the input <i>i</i>-th LSP and the reconstructed (<i>i</i> - 1)-th LSP is computed.</li>
<li>This difference is quantized with a suitable metric, according to the <i>i</i>-th stage quantizer level partition 'seen' by each <i>j</i>-th state.</li>
<li>The Viterbi algorithm is then acted upon on each trellis state, in order to determine the best previous state and, therefore, the updated history path. Furthermore, the quantization accumulated cost is updated for each state; in particular, this cost is consistent with the metric used for each LSP difference quantization.</li>
</ul></p>
<p id="p0034" num="0034">An example of this procedure is depicted in Figure 3, where a simple 4-state trellis with 4 quantization levels (for each quantization step) and 1 bit/branch is used.</p>
<p id="p0035" num="0035">Suppose that the 2nd LSP is input. Hence, the corresponding LSP<!-- EPO <DP n="8"> --> difference must be quantized.</p>
<p id="p0036" num="0036">In the trellis of Figure 3, the quantity <i>D</i><maths id="math0001" num=""><math display="inline"><mrow><mfrac linethickness="0"><mrow><mtext mathvariant="italic">i</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></mfrac></mrow></math><img id="ib0001" file="imgb0001.tif" wi="2" he="9" img-content="math" img-format="tif" inline="yes"/></maths> refers to the quantized <i>i</i>-th LSP difference, according to the corresponding quantization level belonging to the <i>j</i>-th state history path.</p>
<p id="p0037" num="0037">It is clear that by adding all the quantized LSP differences belonging to the generical state history path, the reconstructed 1st LSP may be obtained, as function of the state under consideration.</p>
<p id="p0038" num="0038">In particular, let:<maths id="math0002" num=""><math display="block"><mrow><mtable><mtr><mtd><mrow><mtable><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext>0</mtext></mrow></msub><mtext> = </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>0</mtext></mrow><mrow><mtext>0</mtext></mrow></msubsup><mtext> + </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>0</mtext></mrow><mrow><mtext>1</mtext></mrow></msubsup><mtext>   be the reconstructed LSP along the state 0 path.</mtext></mrow></mtd></mtr><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext>1</mtext></mrow></msub><mtext> = </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>1</mtext></mrow><mrow><mtext>0</mtext></mrow></msubsup><mtext> + </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>1</mtext></mrow><mrow><mtext>1</mtext></mrow></msubsup><mtext>   be the reconstructed LSP along the state 1 path.</mtext></mrow></mtd></mtr><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext>1</mtext></mrow></msub><mtext> = </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>1</mtext></mrow><mrow><mtext>0</mtext></mrow></msubsup><mtext> + </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>1</mtext></mrow><mrow><mtext>1</mtext></mrow></msubsup><mtext>   be the reconstructed LSP along the state 2 path.</mtext></mrow></mtd></mtr><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext>3</mtext></mrow></msub><mtext> = </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>3</mtext></mrow><mrow><mtext>0</mtext></mrow></msubsup><mtext> + </mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>3</mtext></mrow><mrow><mtext>1</mtext></mrow></msubsup><mtext>   be the reconstructed LSP along the state 3 path.</mtext></mrow></mtd></mtr></mtable></mrow></mtd></mtr></mtable></mrow></math><img id="ib0002" file="imgb0002.tif" wi="140" he="32" img-content="math" img-format="tif"/></maths></p>
<p id="p0039" num="0039">For each <i>j</i>-th state the difference between the 2nd input LSP and the reconstructed 1st LSP may be computed. Let this quantity be denoted as <i>D</i><maths id="math0003" num=""><math display="inline"><mrow><mfrac linethickness="0"><mrow><mtext mathvariant="italic">2</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></mfrac></mrow></math><img id="ib0003" file="imgb0003.tif" wi="2" he="9" img-content="math" img-format="tif" inline="yes"/></maths>.</p>
<p id="p0040" num="0040">Afterward, the transition cost from each <i>j</i>-th state to each possible <i>k</i>-th future state is computed. This transition cost is related to the quantization level associated to the corresponding transition branch; with reference to Figure 3, the transition cost is denoted as <i>C</i><sub><i>jk</i></sub>.</p>
<p id="p0041" num="0041">In particular, <i>C</i><sub><i>jk</i></sub> depends on the quantization error which is measured (according to a proper metric) as function of the "transitional" quantization level.</p>
<p id="p0042" num="0042">To be more specific, let <i>L</i><sub><i>jk</i></sub> the quantization level associated to the transition between the <i>j</i>-th state and the <i>k</i>-th one. We can write:<!-- EPO <DP n="9"> -->
<ul id="ul0004" list-style="dash" compact="compact">
<li>Starting from state 0, compute:<maths id="math0004" num=""><math display="block"><mrow><mtable><mtr><mtd><mrow><mtable><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">C</mtext></mrow><mrow><mtext>00</mtext></mrow></msub><mtext> = </mtext><mtext mathvariant="italic">m</mtext><mtext>(</mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>0</mtext></mrow><mrow><mtext>2</mtext></mrow></msubsup><mtext>, </mtext><msub><mrow><mtext mathvariant="italic">L</mtext></mrow><mrow><mtext>00</mtext></mrow></msub><mtext>)   transition cost associated to the state 0 - state 0 branch.</mtext></mrow></mtd></mtr><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">C</mtext></mrow><mrow><mtext>01</mtext></mrow></msub><mtext> = </mtext><mtext mathvariant="italic">m</mtext><mtext>(</mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>0</mtext></mrow><mrow><mtext>2</mtext></mrow></msubsup><mtext>, </mtext><msub><mrow><mtext mathvariant="italic">L</mtext></mrow><mrow><mtext>01</mtext></mrow></msub><mtext>)   transition cost associated to the state 0 - state 1 branch.</mtext></mrow></mtd></mtr></mtable></mrow></mtd></mtr></mtable></mrow></math><img id="ib0004" file="imgb0004.tif" wi="156" he="17" img-content="math" img-format="tif"/></maths></li>
<li>Starting from state 1, compute:<maths id="math0005" num=""><math display="block"><mrow><mtable><mtr><mtd><mrow><mtable><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">C</mtext></mrow><mrow><mtext>12</mtext></mrow></msub><mtext> = </mtext><mtext mathvariant="italic">m</mtext><mtext>(</mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>1</mtext></mrow><mrow><mtext>2</mtext></mrow></msubsup><mtext>, </mtext><msub><mrow><mtext mathvariant="italic">L</mtext></mrow><mrow><mtext>12</mtext></mrow></msub><mtext>)   transition cost associated to the state 1 - state 2 branch.</mtext></mrow></mtd></mtr><mtr><mtd><mrow><msub><mrow><mtext mathvariant="italic">C</mtext></mrow><mrow><mtext>13</mtext></mrow></msub><mtext> = </mtext><mtext mathvariant="italic">m</mtext><mtext>(</mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext>1</mtext></mrow><mrow><mtext>2</mtext></mrow></msubsup><mtext>, </mtext><msub><mrow><mtext mathvariant="italic">L</mtext></mrow><mrow><mtext>13</mtext></mrow></msub><mtext>)   transition cost associated to the state 1 - state 3 branch.</mtext></mrow></mtd></mtr></mtable></mrow></mtd></mtr></mtable></mrow></math><img id="ib0005" file="imgb0005.tif" wi="155" he="17" img-content="math" img-format="tif"/></maths></li>
<li>Similarly, the procedure is repeated for state 2 and state 3. After all the transition costs have been computed, the best state path up to the 3rd trellis stage must be updated (for each trellis state). The well-known Viterbi algorithm is used to this goal. As an example, referring to Figure 3, we have the following:</li>
<li>Current 'observation' state: state 0
<ul id="ul0005" list-style="dash" compact="compact">
<li>Compute the candidate accumulated cost with respect to the previous state 0: <maths id="math0006" num=""><math display="inline"><mrow><mtext mathvariant="italic">A</mtext><mtext> = </mtext><mtext mathvariant="italic">0</mtext><msub><mrow><mtext>(0) + C</mtext></mrow><mrow><mtext>00</mtext></mrow></msub></mrow></math><img id="ib0006" file="imgb0006.tif" wi="29" he="6" img-content="math" img-format="tif" inline="yes"/></maths>, <i>0</i>(0) being the state 0 overall cost accumulated so far.</li>
<li>Compute the candidate accumulated cost with respect to the previous state 2: <maths id="math0007" num=""><math display="inline"><mrow><mtext mathvariant="italic">B</mtext><mtext> = </mtext><mtext mathvariant="italic">0</mtext><mtext>(2) + </mtext><msub><mrow><mtext mathvariant="italic">C</mtext></mrow><mrow><mtext>20</mtext></mrow></msub></mrow></math><img id="ib0007" file="imgb0007.tif" wi="30" he="5" img-content="math" img-format="tif" inline="yes"/></maths>, <i>0</i>(2) being the state 2 overall cost accumulated so far.</li>
<li>If <i>A</i> &lt; <i>B</i> the state 0 is considered to be the best previous state with respect to the current state 0; the new state 0 path is<!-- EPO <DP n="10"> --> determined by concatenating the state 0 path with the state 0 - state 0 transition. <br/>
On the contrary, if <i>A</i> &gt; <i>B</i>, the state 0 path is updated according to the state 2 path and to the state 2 - state 0 transition.
</li>
</ul></li>
<li>The same procedure applies for the 'observation' states 2, 3, 4.</li>
</ul></p>
<p id="p0043" num="0043">Finally, suppose that the following situation occurs:
<ul id="ul0006" list-style="dash" compact="compact">
<li>Best previous state with respect to state 0: <i>state</i>0</li>
<li>Best previous state with respect to state 1: <i>state</i>2</li>
<li>Best previous state with respect to state 2: <i>state</i>3</li>
<li>Best previous state with respect to state 3: <i>state</i>3</li>
</ul></p>
<p id="p0044" num="0044">The state paths are updated as depicted in Figure 4. Furthermore, the overall accumulated costs are updated for each trellis state.</p>
<p id="p0045" num="0045">We are ready for the next LSP quantization.</p>
<p id="p0046" num="0046">After the <i>p</i>-th LSP difference (p being the predictor order) has been quantized, the final state with the minimum accumulated cost is selected as the "winning" one. Its index (or, equivalently, the index of the corresponding initial state) is transmitted, together with the state transition labels of its history path.</p>
<p id="p0047" num="0047">At the decoder, the initial winning state index and the state transition labels of its history path are input. All the LSP differences can be recovered from the state transition label pointers to the quantization level table. Afterward, the LSP frequencies can be reconstructed.</p>
<p id="p0048" num="0048">Besides an intra-frame correlation of the LSP parameters (i.e. the ordering property), it is possible to take advantage of the strong<!-- EPO <DP n="11"> --> inter-frame correlation; this may lead to efficient quantization schemes that can operate, for instance, in a two-dimensional differential fashion. A possible application is described in C.C. Kuo, F.R. Jean, H.C. Wang, "Low Bit-Rate Quantization of LSP Parameters Using Two-Dimensional Differential Coding", Proc. ICASSP '92, p. 97-100, where a two-dimensional differential coding scheme is shown to significantly improve the effectiveness of the quantization scheme as function of the desired encoding rate.</p>
<p id="p0049" num="0049">In particular, in C.C. Kuo, F.R. Jean, H.C. Wang, "Low Bit-Rate Quantization of LSP Parameters Using Two-Dimensional Differential Coding", Proc. ICASSP '92, pagg. 97-100, the inter-frame/intra-frame dependency of the <i>i</i>-th LSP at frame n may be expressed as:<maths id="math0008" num="(1)"><math display="block"><mrow><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><mtext mathvariant="italic">(n)</mtext><mtext> = </mtext><msub><mrow><mtext mathvariant="italic">a</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) + </mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><mtext mathvariant="italic">(n</mtext><mtext> - 1),</mtext><mtext mathvariant="italic">i</mtext><mtext> = 1,2,...</mtext><mtext mathvariant="italic">p</mtext></mrow></math><img id="ib0008" file="imgb0008.tif" wi="74" he="6" img-content="math" img-format="tif"/></maths> where <i>p</i> is the predictor order and <i>a</i><sub><i>i</i></sub> and <i>b</i><sub><i>i</i></sub> are the coefficients of an optimal two-dimensional (i.e. 2 - <i>D</i>) predictor; these coefficients can be estimated from a long sequence of speech, as described in C.C. Kuo, F.R. Jean, H.C. Wang, "Low Bit-Rate Quantization of LSP Parameters Using Two-Dimensional Differential Coding", Proc. ICASSP '92, pagg. 97-100. <i>f</i><sub><i>i</i></sub>(<i>n</i>) is the current LSP estimation; <i>f</i><sub><i>i</i></sub><sub>-1</sub>(<i>n</i>) and <i>f</i><sub><i>i</i></sub>(<i>n</i>-1) are previously quantized parameters.</p>
<p id="p0050" num="0050">Therefore, it is possible to transmit the difference between the exact value and the estimated value of the current LSP.</p>
<p id="p0051" num="0051">It is clear that this quantization scheme may not be the optimal one<!-- EPO <DP n="12"> --> in case channel errors occur during the transmission of the LSP difference information.</p>
<p id="p0052" num="0052">It is possible to cope with this problem by employing a careful design of the 2-D predictor coefficients. However; this issue will be described in greater details in a next paragraph. For the time being, we will assume to deal with proper 2-D predictor coefficients and will concentrate on the TCQ functioning in this case.</p>
<p id="p0053" num="0053">The working principle is analogous to the one described previously for the 1-D (i.e. intra-frame) case, which simply exploits the LSP,ordering property. In particular, it is worth to note that the 1-D case can be considered as a particular case of the 2-D scheme, by putting <i>a</i><sub><i>i</i></sub> = 1 and <i>b</i><sub><i>i</i></sub> = 0.</p>
<p id="p0054" num="0054">In particular, suppose that the <i>i</i>-th LSP 2-D difference must be quantized. The following operations will be performed:</p>
<p id="p0055" num="0055">For each <i>j</i>-th trellis state of the (<i>i</i> - 1)-th trellis stage, a corresponding (reconstructed) <i>LSP</i><maths id="math0009" num=""><math display="inline"><mrow><mfrac linethickness="0"><mrow><mtext>i-</mtext></mrow><mrow><mtext>j</mtext></mrow></mfrac></mrow></math><img id="ib0009" file="imgb0009.tif" wi="2" he="9" img-content="math" img-format="tif" inline="yes"/></maths>  (<i>n</i>) will be available (<i>n</i> is the current frame index).
<ul id="ul0007" list-style="dash" compact="compact">
<li>For each <i>j</i>-th state, the 2-D LSP difference is computed between the input <i>i</i>-th LSP and the weighted combination of:
<ul id="ul0008" list-style="dash" compact="compact">
<li>the reconstructed <i>LSP</i><maths id="math0010" num=""><math display="inline"><mrow><mfrac linethickness="0"><mrow><mtext>i-</mtext></mrow><mrow><mtext>j</mtext></mrow></mfrac></mrow></math><img id="ib0010" file="imgb0010.tif" wi="2" he="9" img-content="math" img-format="tif" inline="yes"/></maths></li>
<li>the corresponding <i>i</i>-th quantized LSP derived in the previous frame <i>LSP</i><sup><i>i</i></sup>(<i>n</i>-1).</li>
</ul> The combination weights are the 2-D predictor coefficients.</li>
<li>The computed difference is then quantized according to the trellis<!-- EPO <DP n="13"> --> topology and to the quantization level configuration. The quantization procedure is analogous to the one described in the 1-D case.</li>
<li>Once all the LSP differences have been quantized and the state paths (as well as their accumulated costs) have been updated, the quantized <i>LSP</i><maths id="math0011" num=""><math display="inline"><mrow><mfrac linethickness="0"><mrow><mtext mathvariant="italic">i</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></mfrac></mrow></math><img id="ib0011" file="imgb0011.tif" wi="2" he="9" img-content="math" img-format="tif" inline="yes"/></maths> (n) can be derived, as function of each <i>j</i>-th state under consideration. In particular, we have:<maths id="math0012" num="(2)"><math display="block"><mrow><msubsup><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msubsup><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) = </mtext><msub><mrow><mtext mathvariant="italic">a</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msup><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msup><msup><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext mathvariant="italic">previous</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>(</mtext><mtext mathvariant="italic">j</mtext><mtext>)</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) + </mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msup><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msup><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext> - 1) + </mtext><mtext mathvariant="italic">D</mtext><msubsup><mrow><mtext>​</mtext></mrow><mrow><msub><mrow><mtext mathvariant="italic">LSP</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msubsup><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>),</mtext></mrow></math><img id="ib0012" file="imgb0012.tif" wi="109" he="8" img-content="math" img-format="tif"/></maths> where<img id="ib0013" file="imgb0013.tif" wi="13" he="10" img-content="undefined" img-format="tif"/> is the best local quantized LSP difference appertaining to the <i>j</i>-th state and <i>previous (j)</i> is its previous state in the history path, as derived from the Viterbi algorithm.</li>
</ul></p>
<p id="p0056" num="0056">Iterating this way, all the LSP can be quantized, according to the 2-D predictor behaviour.</p>
<p id="p0057" num="0057">More in general, a multi-coefficient 2-D predictor can be employed, both in the intra-frame and in the inter-frame sense, as follows:<maths id="math0013" num="(3)"><math display="block"><mrow><apply><sum/><lowlimit><mtext mathvariant="italic">j</mtext><mtext>=0</mtext></lowlimit><uplimit><mtext mathvariant="italic">J</mtext></uplimit><mrow><mtext> </mtext><msub><mrow><mtext mathvariant="italic">a</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext><mtext>-</mtext><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>)</mtext></mrow></apply><mtext>+</mtext><apply><sum/><lowlimit><mtext mathvariant="italic">k</mtext><mtext>=0</mtext></lowlimit><uplimit><mtext mathvariant="italic">K</mtext></uplimit><mrow><mtext> </mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">k</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>-</mtext><mtext mathvariant="italic">k</mtext><mtext>)</mtext></mrow></apply></mrow></math><img id="ib0014" file="imgb0014.tif" wi="59" he="9" img-content="math" img-format="tif"/></maths></p>
<p id="p0058" num="0058">It is worth to note that the predictor length is not necessarily the same in each dimension.</p>
<p id="p0059" num="0059">A further way to exploit the 'spatial-temporal' redundancy of the LSP parameters is to use a 3-D predictor as follows:<maths id="math0014" num="(4)"><math display="block"><mrow><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) = </mtext><msub><mrow><mtext mathvariant="italic">a</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) + </mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>-1) + </mtext><msub><mrow><mtext mathvariant="italic">c</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">i</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>-1)</mtext></mrow></math><img id="ib0015" file="imgb0015.tif" wi="72" he="6" img-content="math" img-format="tif"/></maths> that is, we introduce another inter-frame/intra-frame dependency,<!-- EPO <DP n="14"> --> namely the one related to the previous (in the intra-frame sense) LSP of the previous (in the inter-frame sense) frame. The third weighting coefficient can be determined in an 'optimal' way, as will be described in a following section.</p>
<p id="p0060" num="0060">The concept can be extended further, by introducing a multi-coefficient multi-dimensional predictor, operating with different prediction orders, according to the prediction 'direction' (i.e. intra-frame, inter-frame, various intra-frame/inter-frame combinations).</p>
<p id="p0061" num="0061">Irrespectively of the predictor structure, a difference between a LSP and the corresponding estimated one will be quantized, following the trellis search procedure and the Viterbi algorithm described previously.</p>
<p id="p0062" num="0062">At the decoder site, all the LSP differences can be recovered from the best state information and the related history path. The LSP values can then be reconstructed by re-adding the previously (both in the intra-frame and in the inter-frame sense) reconstructed parameters, after weighting them by the corresponding predictor coefficients.</p>
<p id="p0063" num="0063">It is clear that the TCQ of the LSP parameters (in a generical differential sense) can be carried out according to any suitable metric that allows to measure an overall distortion as function of successive local distortions.</p>
<p id="p0064" num="0064">In particular, a simple mean squared error (MSE) could be used as the local metric for the quantization error. In this respect, the<!-- EPO <DP n="15"> --> transition cost defined previously (i.e. see the 1-D case) could be defined as:<maths id="math0015" num="(5)"><math display="block"><mrow><msub><mrow><mtext mathvariant="italic">C</mtext></mrow><mrow><mtext mathvariant="italic">jk</mtext></mrow></msub><mtext> = (</mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow><mrow><mtext mathvariant="italic">k</mtext></mrow></msubsup><mtext> - </mtext><msub><mrow><mtext mathvariant="italic">L</mtext></mrow><mrow><mtext mathvariant="italic">jk</mtext></mrow></msub><msup><mrow><mtext>)</mtext></mrow><mrow><mtext>2</mtext></mrow></msup></mrow></math><img id="ib0016" file="imgb0016.tif" wi="34" he="7" img-content="math" img-format="tif"/></maths></p>
<p id="p0065" num="0065">More in general, a weighted MSE (WMSE) could be employed, following (e.g.) the guidelines specified in K.K. Paliwal, B.S. Atal, "Efficient Vector Quantization of LPC Parameters at 24 Bits/Frame", Proc. ICASSP '91, p. 661-664, where the spectral content of the speech signal at the LSP frequency location is taken into account explicitly. Or, a WMSE criterion that considers the relative weight of the specific LSP that is being quantized could also be considered. In this case, formula (5) could be re-written as:<maths id="math0016" num="(6)"><math display="block"><mrow><msub><mrow><mtext mathvariant="italic">C</mtext></mrow><mrow><mtext mathvariant="italic">jk</mtext></mrow></msub><mtext> = </mtext><mtext mathvariant="italic">f</mtext><mtext>(</mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow><mrow><mtext mathvariant="italic">k</mtext></mrow></msubsup><mtext>,</mtext><msub><mrow><mtext mathvariant="italic">L</mtext></mrow><mrow><mtext mathvariant="italic">jk</mtext></mrow></msub><mtext>)(</mtext><msubsup><mrow><mtext mathvariant="italic">D</mtext></mrow><mrow><mtext mathvariant="italic">J</mtext></mrow><mrow><mtext mathvariant="italic">k</mtext></mrow></msubsup><mtext> - </mtext><msub><mrow><mtext mathvariant="italic">L</mtext></mrow><mrow><mtext mathvariant="italic">jk</mtext></mrow></msub><msup><mrow><mtext>)</mtext></mrow><mrow><mtext>2</mtext></mrow></msup></mrow></math><img id="ib0017" file="imgb0017.tif" wi="51" he="7" img-content="math" img-format="tif"/></maths> where <i>f</i>(.,.) is a (one-dimensional or two-dimensional) weighting function that would take into account the differential LSP to be quantized and/or the quantization level that is being considered.</p>
<p id="p0066" num="0066">Although LSP have proved to be a useful representation of the LPC coefficients with respect to quantization effectiveness, the reflection coefficients are also attractive, for some reasons like:
<ul id="ul0009" list-style="dash" compact="compact">
<li>Easy control of the filter stability.</li>
<li>No need of complicated arithmetic procedures to convert them into LSP parameters.</li>
<li>Possibility of implementing the necessary filter structures in lattice forms, with evident advantages for fix-point computation.</li>
</ul></p>
<p id="p0067" num="0067">A recursive structure may be used for the computation of the<!-- EPO <DP n="16"> --> reflection coefficients, starting either from the values of the autocorrelation function (and thereby using the well-known Leroux-Gueguen algorithm) or from the values of the signal covariance function (by employing the so-called covariance-lattice formulation, as explained in A. Cumani, "On a Covariance-Lattice Algorithm for Linear Prediction", Proc. ICASSP '82, pagg. 651-654).</p>
<p id="p0068" num="0068">In particular, the Leroux-Gueguen algorithm should be reformulated properly in order to take into account the eventual quantization of the reflection coefficients after their computation at each step of the recursion.</p>
<p id="p0069" num="0069">This gives rise to a slightly modified recursive algorithm in which, starting from the autocorrelation values, the reflection coefficients are computed as follows:
<ul id="ul0010" list-style="dash" compact="compact">
<li>Let <i>f</i><sub><i>i</i></sub>(<i>n</i>) be the forward residual of the lattice structure <i>j</i>-th stage and b<sub><i>j</i></sub>(<i>n</i>) be the corresponding backward residual. Then, the expression of the <i>j</i>-th stage residuals as function of the (<i>j</i> - 1)-th stage ones is as follows:<maths id="math0017" num=""><math display="block"><mrow><mtable><mtr><mtd><mrow><mtable><mtr><mtd><mrow><mtable><mlabeledtr><mtext>(7)</mtext><mtd><mrow><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) = </mtext><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) + </mtext><msub><mrow><mtext mathvariant="italic">K</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext> - 1)</mtext></mrow></mtd></mlabeledtr></mtable></mrow></mtd></mtr><mtr><mtd><mrow><mtable><mlabeledtr><mtext>(8)</mtext><mtd><mrow><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) = </mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">j-</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>1</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext> - 1) + </mtext><msub><mrow><mtext mathvariant="italic">K</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>), </mtext></mrow></mtd></mlabeledtr></mtable></mrow></mtd></mtr></mtable></mrow></mtd></mtr></mtable></mrow></math><img id="ib0018" file="imgb0018.tif" wi="183" he="21" img-content="math" img-format="tif"/></maths> <i>K</i><sub><i>j</i></sub> being the <i>j</i>-th stage reflection coefficients.</li>
<li>Defining the initial conditions:<maths id="math0018" num="(9)"><math display="block"><mrow><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext>0</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) = </mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext>0</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>) + </mtext><mtext mathvariant="italic">s</mtext><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>)</mtext></mrow></math><img id="ib0019" file="imgb0019.tif" wi="38" he="6" img-content="math" img-format="tif"/></maths> <i>s(n)</i> being the lattice structure input signal.</li>
<li>Defining also the following autocorrelation and cross-correlation<!-- EPO <DP n="17"> --> functions:<maths id="math0019" num="(10)"><math display="block"><mrow><msubsup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">F</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msubsup><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext>) = </mtext><apply><sum/><lowlimit><mtext mathvariant="italic">n</mtext><mtext>=0</mtext></lowlimit><uplimit><mtext mathvariant="italic">N</mtext><mtext>-1-</mtext><mtext mathvariant="italic">k</mtext></uplimit><mrow><mtext> </mtext><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>)</mtext><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>+</mtext><mtext mathvariant="italic">k</mtext><mtext>)</mtext></mrow></apply></mrow></math><img id="ib0020" file="imgb0020.tif" wi="55" he="9" img-content="math" img-format="tif"/></maths> autocorrelation of the forward residual at the <i>j</i>-th lattice stage.<maths id="math0020" num="(11)"><math display="block"><mrow><msubsup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">B</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msubsup><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext>) = </mtext><apply><sum/><lowlimit><mtext mathvariant="italic">n</mtext><mtext>=0</mtext></lowlimit><uplimit><mtext mathvariant="italic">N</mtext><mtext>-1-</mtext><mtext mathvariant="italic">k</mtext></uplimit><mrow><mtext> </mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>)</mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>+</mtext><mtext mathvariant="italic">k</mtext><mtext>)</mtext></mrow></apply></mrow></math><img id="ib0021" file="imgb0021.tif" wi="57" he="8" img-content="math" img-format="tif"/></maths> autocorrelation of the backward residual at the <i>j</i>-th lattice stage.<maths id="math0021" num="(12)"><math display="block"><mrow><msubsup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">FB</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msubsup><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext>) = </mtext><apply><sum/><lowlimit><mtext mathvariant="italic">n</mtext><mtext>=0</mtext></lowlimit><uplimit><mtext mathvariant="italic">N</mtext><mtext>-1-</mtext><mtext mathvariant="italic">k</mtext></uplimit><mrow><mtext> </mtext><msub><mrow><mtext mathvariant="italic">f</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>)</mtext><msub><mrow><mtext mathvariant="italic">b</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext> + </mtext><mtext mathvariant="italic">k</mtext><mtext> - 1),</mtext></mrow></apply></mrow></math><img id="ib0022" file="imgb0022.tif" wi="69" he="9" img-content="math" img-format="tif"/></maths> cross-correlation between the forward and backward residuals at the <i>j</i>-th lattice stage.</li>
<li>The forward residual autocorrelation at the <i>j</i>-th lattice stage may be expressed by means of the following recursive formula:<maths id="math0022" num="(13)"><math display="block"><mrow><msubsup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">F</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msubsup><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext>) = </mtext><msup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msup><msup><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext mathvariant="italic">F</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext>) + </mtext><msub><mrow><mtext mathvariant="italic">K</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msup><msup><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext mathvariant="italic">BF</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext> + 1) + </mtext><msub><mrow><mtext mathvariant="italic">K</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msup><msup><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext mathvariant="italic">FB</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext> - 1) + </mtext><msup><mrow><mtext mathvariant="italic">K</mtext></mrow><mrow><mtext>2</mtext></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msub><msup><mrow><mtext mathvariant="italic">R</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msup><msup><mrow><mtext>​</mtext></mrow><mrow><mtext>-1</mtext></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext mathvariant="italic">B</mtext></mrow></msub><mtext>(</mtext><mtext mathvariant="italic">k</mtext><mtext>)</mtext></mrow></math><img id="ib0023" file="imgb0023.tif" wi="130" he="7" img-content="math" img-format="tif"/></maths></li>
<li>Therefore, the optimal value for the reflection coefficient <i>K</i><sub><i>j</i></sub> is the one that minimizes the forward residual energy <i>R</i><sub><i>F</i></sub>(0).</li>
<li>Once the <i>K</i><sub><i>j</i></sub> values is computed (and, eventually, quantized), the autocorrelation may be updated, as well as the quantities, <i>R</i><sup><i>j</i></sup><sub><i>B</i></sub>, <i>R</i><sup><i>i</i></sup><sub><i>FB</i></sub>, <i>R</i><sup><i>i</i></sup><sub><i>BF</i></sub> (using expressions similar to the one derived for <i>R</i><sup><i>i</i></sup><sub><i>F</i></sub>)</li>
</ul></p>
<p id="p0070" num="0070">Both the covariance-lattice formulation and the modified autocorrelation one are particularly amenable to TCQ, in that each computed reflection coefficient can be quantized prior to the next recursion step.</p>
<p id="p0071" num="0071">Again, after defining the trellis topology and the quantization level number at each quantization step, each reflection coefficient can<!-- EPO <DP n="18"> --> be computed as function of each particular state.</p>
<p id="p0072" num="0072">Its value can therefore take into account the quantization error accumulated along each branch of a generical state path.</p>
<p id="p0073" num="0073">Afterwards, the computed reflection coefficient can be quantized according to the quantization level subset 'seen' by each particular trellis state.</p>
<p id="p0074" num="0074">In formulas, the recursive algorithm for reflection coefficient computation, with embedded TCQ may be stated as follows (only the formulation related to the covariance-lattice approach is given, since the corresponding formalism for the autocorrelation approach may be derived in an analogous way; also, note that the formalism used resembles the one described in: A. Cumani, "On a Covariance-Lattice Algorithm for Linear Prediction", Proc. ICASSP '82, pagg. 651-654.
<ul id="ul0011" list-style="dash" compact="compact">
<li>Given a block of N signal samples: <i>s</i>(0), <i>s</i>(1),..., <i>s</i>(<i>N</i> - 1), compute the covariance Φ<sub><i>ik</i></sub>, for <i>i</i>, <i>k</i> = 0,1,...,<i>p</i> (<i>p</i> being the predictor order):<maths id="math0023" num="(14)"><math display="block"><mrow><msub><mrow><mtext>Φ</mtext></mrow><mrow><mtext mathvariant="italic">ik</mtext></mrow></msub><mtext> = </mtext><apply><sum/><lowlimit><mtext mathvariant="italic">n</mtext><mtext>=</mtext><mtext mathvariant="italic">p</mtext></lowlimit><uplimit><mtext mathvariant="italic">N</mtext><mtext>-1</mtext></uplimit><mrow><mtext> </mtext><mtext mathvariant="italic">s</mtext><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>-1)</mtext><mtext mathvariant="italic">s</mtext><mtext>(</mtext><mtext mathvariant="italic">n</mtext><mtext>-</mtext><mtext mathvariant="italic">k</mtext><mtext>)</mtext></mrow></apply></mrow></math><img id="ib0024" file="imgb0024.tif" wi="48" he="8" img-content="math" img-format="tif"/></maths></li>
<li>Set up <i>F</i><sup><i>0</i></sup><sub><i>ij</i></sub>, <i>B</i><sup><i>0</i></sup><sub><i>ij</i></sub>, <i>C</i><sup><i>0</i></sup><sub><i>ij</i></sub>, for <i>i</i>,<i>j</i> = 0,1,...,<i>p</i> - 1 using formula (14) of the reference mentioned above. Also, set: <i>m</i> = 0</li>
<li>For each predictor stage m:
<ul id="ul0012" list-style="dash" compact="compact">
<li>Compute the <i>m</i>-th reflection coefficient as function of the <i>j</i>-th TCQ state:<br/>
where <i>C</i><sub>00</sub><sup><i>mj</i></sup>, <i>F</i><sub>00</sub><sup><i>mj</i></sup> and <i>B</i><sub>00</sub><sup><i>mj</i></sup> are the 'forward-backward covariance<!-- EPO <DP n="19"> --><maths id="math0024" num=""><math display="block"><mrow><msubsup><mrow><mtext mathvariant="italic">K</mtext></mrow><mrow><mtext mathvariant="italic">m</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msubsup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>+1</mtext></mrow></msub><mtext> = -2</mtext><mtext mathvariant="italic">C</mtext><msup><mrow><mtext>​</mtext></mrow><mrow><msup><mrow><mtext mathvariant="italic">m</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msup></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>00</mtext></mrow></msub><mtext>/(</mtext><mtext mathvariant="italic">F</mtext><msup><mrow><mtext>​</mtext></mrow><mrow><msup><mrow><mtext mathvariant="italic">m</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msup></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>00</mtext></mrow></msub><mtext> + </mtext><mtext mathvariant="italic">B</mtext><msup><mrow><mtext>​</mtext></mrow><mrow><msup><mrow><mtext mathvariant="italic">m</mtext></mrow><mrow><mtext mathvariant="italic">j</mtext></mrow></msup></mrow></msup><msub><mrow><mtext>​</mtext></mrow><mrow><mtext>00</mtext></mrow></msub><mtext>)</mtext></mrow></math><img id="ib0025" file="imgb0025.tif" wi="62" he="7" img-content="math" img-format="tif"/></maths> functions' that, once determined making use of formulas (12a,b,c) of A. Cumani, "On a Covariance-Lattice Algorithm for Linear Prediction", Proc. ICASSP '82, pagg. 651-654, take into account the quantization values of the previous reflection coefficients along the <i>j</i>-th TCQ state path.</li>
<li>Quantize the reflection coefficient just computed according to the quantizer level partition appertaining to the <i>j</i>-th state. In particular, a non-linear transformation (e.g. log-area ratios) can be done prior to quantization. <br/>
Each quantization level 'seen' by the <i>j</i>-th state will correspond to a particular state branch connecting the <i>j</i>-th state to a future <i>i</i>-th state (according to the trellis topology).
</li>
<li>The optimal local quantization level thus found is related to a local transition cost, between the <i>j</i>-th state and the <i>i</i>-th one.</li>
<li>Proceed to the next trellis stage and update the accumulated cost of each state (making use of the accumulated cost of the previous trellis stage and of the local metrics just computed). Also, update the partial quantization path for each state. That is, for each <i>i</i>-th trellis state update the 'forward-backward covariance functions' making use of the function values in the previous <i>j</i>-th state and the quantization level of the reflection coefficient that corresponds to the <i>j</i>-th state -- <i>i</i>-th state transition branch.</li>
</ul><!-- EPO <DP n="20"> --></li>
<li>Again, the state with the minimum overall accumulated cost is declared as 'winner'. Its value, together with the trellis labels defining its path, determines the quantized reflection coefficient vector.</li>
</ul></p>
<p id="p0075" num="0075">As for the LSP case, a proper metric should be employed to carry out the quantization process; in particular, a matric that allows to measure an overall distortion as function of successive local distortions (such as a MSE- or WMSE-based metric) can be suitable.</p>
<p id="p0076" num="0076">The quantization procedures outlined in the previous paragraphs may not be the optimal ones (with respect to both the LSP and the reflection coefficients).</p>
<p id="p0077" num="0077">In particular, the importance of using suitable metrics has been stressed, that allow the computation of the accumulated cost as sum of successive partial costs.</p>
<p id="p0078" num="0078">From the perceptual point of view, it is well known that the most reliable metric for measuring the effectiveness of the LPC parameter quantization is based on the cepstrual coefficients (e.g. see K.K. Paliwal, B.S. Atal, "Efficient Vector Quantization of LPC Parameters at 24 Bits/Frame", Proc. ICASSP '91, pagg. 661-664). In particular, once two sets of predictor coefficients (i.e. before and after quantization) are known, one should compute the corresponding cepstrual coefficient sequences and then measure the MSE (namely, the cepstral distance CD).</p>
<p id="p0079" num="0079">However, this procedure is not feasible in carrying out the<!-- EPO <DP n="21"> --> step-by-step quantization (i.e. as a new LPC parameter is available), as described in the previous TCQ procedures.</p>
<p id="p0080" num="0080">Therefore, in order to obtain the set of quantized coefficients that guarantee the best perceptual LPC reproduction, the following steps should carried out:
<ul id="ul0013" list-style="dash" compact="compact">
<li>Reconstruct the decoding path in correspondence of each trellis state, therefore obtaining a set of quantized LPC parameter vectors (i.e. either in terms of LSP or in terms of reflection coefficients).</li>
<li>For each vector, obtain the corresponding representation in terms of LPC cepstrual coefficients and measure the CD with respect to the cepstral coefficient representation of the unquantized model.</li>
<li>The trellis parameters to be transmitted should be the ones that define the best LPC vector (in terms of cepstral distance).</li>
</ul></p>
<p id="p0081" num="0081">Obtaining the cepstral coefficients from a set of LPC parameters (i.e. LSP or reflection coefficients) is not a trivial task. Therefore, the outlined procedure is likely to be very time-consuming. It is possible to reduce the computation load by reconstructing the decoding path in correspondence of only a subset of the overall trellis states.</p>
<p id="p0082" num="0082">This implies that an implicit assumption is made: namely, the trellis state subset with lowest overall accumulated cost is likely to contain the best trellis state, in terms of CD.</p>
<p id="p0083" num="0083">For better quantization efficiency, the quantization levels are different for each trellis state; the following example clarifies this concept (assume that we are dealing with the quantization of the<!-- EPO <DP n="22"> --> LSP parameters.</p>
<p id="p0084" num="0084">The same rationales apply for the quantization of the reflection coefficients as well).</p>
<p id="p0085" num="0085">Suppose that the i-th LSP difference must be quantized; its value is computed by taking the difference between the i-th LSF and the reconstructed (i.e. quantized) (<i>i</i> - 1)-th LSF. This reconstructed (<i>i</i> - 1)-th LSF is different for each trellis state; the <i>i</i>-th LSF difference thus obtained must be quantized accordingly to the level partition "seen" by the corresponding trellis state.</p>
<p id="p0086" num="0086">Assume that two generical trellis states point to the same subset of quantization level; in standard TCQ procedures (e.g. see M.W. Marcellin, T.G. Fisher, "Trellis Coded Quantization of Memoryless and Gauss-Markov Sources", IEEE Trans. on Communication, vol. 38, No. 1, January 1990) the subset quantization values are the same for the two states; they are only addressed in a different way.</p>
<p id="p0087" num="0087">On the contrary we use different values in the same quantization level subset, as function of TCQ state under consideration.</p>
<p id="p0088" num="0088">In order to obtain this, a proper TCQ training procedure can be adopted; in particular we start from a unique set of quantization values for each state subset; these values can be found by using a standard scalar quantization clustering procedure.</p>
<p id="p0089" num="0089">Afterwards an iterative procedure is adopted in which a long training sequence of LSP is input to the TCQ and the input LSP vector is then assigned to the "partition" corresponding to the obtained TCQ path.<!-- EPO <DP n="23"> --></p>
<p id="p0090" num="0090">At the end of the training procedure each possible quantization path is assigned a partition; the corresponding "cluster vector" can be derived by simply taking a proper mean of each partition value and assigning this mean value to the corresponding path state.</p>
<p id="p0091" num="0091">Next, the LSF value training sequence is again input to the TCQ; a new partition set can be generated and the corresponding set of cluster vectors can be found.</p>
<p id="p0092" num="0092">More in details, during a generical iteration step, the following operations can be performed:
<ul id="ul0014" list-style="dash" compact="compact">
<li>Reset all the partitions belonging to each trellis state. Note that the number of partitions appertaining to each trellis state is equal to the number of quantization levels that can be 'reached' from the state.</li>
<li>For each LSP input vector:
<ul id="ul0015" list-style="dash" compact="compact">
<li>Find the optimal quantized vector according to a predefined metric. The quantized vector is identified by specifying the winning starting state and the branch labels along the state path.</li>
<li>Assign each <i>j</i>-th element of the input vector to a partition whose index corresponds to the branch label of the <i>j</i>-th quantization step along the winning state path. In particular, if the simple MSE is adopted in the quantization phase, the <i>j</i>-th input vector element is simply added to the previous partition value.</li>
</ul></li>
<li>At this point, all the partitions for each trellis state and for each quantization step have been constructed. Each partition 'centroid'<!-- EPO <DP n="24"> --> can be recomputed by taking an appropriate mean as function of the accumulated partition value and of the number of elements inside the partition. In particular, if the MSE metric is adopted in the quantization phase, each centroid can be computed by taking the simple arithmetic mean of the partition accumulated value.</li>
</ul></p>
<p id="p0093" num="0093">Iterating this way it can be observed that the quantization error (in a MSE sense, or in a WMSE sense, according to the metric adopted) is decreasing; although this may not correspond to a performance increase in the cepstral distance sense, it is possible to run the iterative procedure for a fixed iteration number and then choose the quantization level set that guarantees the best performance in terms of cepstral distance.</p>
<p id="p0094" num="0094">Finally, it is worth to note that this iterative reoptimization of the quantization levels is independent on the trellis topology as well as on the LPC parameters under consideration (i.e. LSP or reflection coefficients). However, care must be taken in the determination of the partition centroid, according to the metric used in the quantization phase.</p>
<p id="p0095" num="0095">Trellis Coded Vector Quantization (TCVQ) is a generalization of the TCQ concept. It has been introduced in T.G. Fisher, M.W. Marcellin, M.Wang, "Trellis Coded Vector Quantization", IEEE Trans. on Information Theory, Vol. IT-37, Nov. 1991. and, again, consists of using a structured codebook with an expanded set of quantization levels.</p>
<p id="p0096" num="0096">In particular, instead of dealing with scalar quantization levels, we<!-- EPO <DP n="25"> --> have an expanded set of reproduction vectors. Again, the trellis structure prunes the expanded number of quantization reproduction vectors down to the desired encoding rate.</p>
<p id="p0097" num="0097">When applied to the LPC parameter quantization, the same strategies can be employed, whether we use the representation in terms of LSP or in terms of reflection coefficients.</p>
<p id="p0098" num="0098">It is clear that, for typical predictor orders (i.e. 10), it is not worth to use high-dimension vectors, in order to maintain a favourable trade-off between performance and encoding rate.</p>
<p id="p0099" num="0099">To be more specific, let's consider the following example:
<ul id="ul0016" list-style="dash" compact="compact">
<li>Predictor order = 10 (i.e. 10 LPC coefficients to be quantized)</li>
<li>Scalar quantization versus TCQ case
<ul id="ul0017" list-style="dash" compact="compact">
<li>Using 3 bits/coefficient (8 quantization levels), scalar quantization implies quantizing the LPC information with 30 bits.</li>
<li>Using 2 bits/coefficients (8 quantization levels) and a 16-state trellis (which is a good compromise between performance and computation load), TCQ implies quantizing the LPC information with 4 + 20 = 24 bits</li>
</ul></li>
<li>Vector quantization versus TCVQ case
<ul id="ul0018" list-style="dash" compact="compact">
<li>Dividing the LPC vector in two subvectors of 5 coefficients each and quantizing each subvector with a 2<sub>15</sub>-element codebook (in order to maintain the same encoding rate as for the scalar quantization case), and using, again, a 16-state trellis, the TCVQ approach allows for an overall encoding rate of 4+14+14 = 32<!-- EPO <DP n="26"> --> bits for the LPC information (note that by using simple VQ, we would obtain an encoding rate of 30 bits).</li>
<li>Dividing the LPC vector in 5 subvectors of 2 coefficients each and quantizing each subvector with a 2<sub>6</sub>-element codebook (thus obtaining again an encoding rate of 30 bits for the VQ case), the TCVQ approach with a 16-state trellis would allow to obtain 4+5*5 = 29 bits.</li>
</ul></li>
</ul></p>
<p id="p0100" num="0100">Actually, the TCVQ technique acted upon subvectors of coefficient couples seems a good compromise between encoding rate and performance.</p>
<p id="p0101" num="0101">The TCVQ procedure is carried out in exactly the same way as for the TCQ counterpart, both in the 1-D case (i.e., taking into account only the intra-frame dependency of the LSP parameters), and in the case of multi-dimensional prediction (i.e., exploiting both the intra-frame and the inter-frame dependency of LSP parameters, with any prediction length in either direction).</p>
<p id="p0102" num="0102">Besides, when considering the reflection coefficient case (where the prediction-based solution may not be the optimal one), the TCVQ procedure can be carried out by recursive quantization of reflection coefficient couples (if the subvector dimension is actually 2), by using the same strategy employed for the TCQ case.</p>
<p id="p0103" num="0103">To be more specific, a brief description of the TCVQ procedure, when applied to the LSP in the 1-D case, is as follows (assuming to deal with subvectors of coefficient couples. Also, suppose we want to quantize the successive LSP differences): As an example,<!-- EPO <DP n="27"> --> suppose that the <i>i</i>-th LSP and the (<i>i</i> + 1)-th one must be quantized.</p>
<p id="p0104" num="0104">The following operations will be performed:
<ul id="ul0019" list-style="dash" compact="compact">
<li>For each <i>j</i>-th state of the (<i>i</i> - 1)-th trellis stage, the corresponding (<i>i</i> - 1)-th LSP is reconstructed, by adding all the quantized LSP difference couples belonging to the <i>j</i>-th state history path.</li>
<li>For each <i>j</i>-th state, the LSP difference between the input <i>i</i>-th LSP and the reconstructed (<i>i</i> - 1)-th LSP is computed. Furthermore, the LSP difference between the two input LSP parameters can be computed. This gives rise to a 2-component LSP difference vector to be quantized.</li>
<li>This difference vector is quantized with a suitable metric, according to the <i>i</i>-th stage quantizer level partition 'seen' by each <i>j</i>-th state.</li>
<li>The Viterbi algorithm is then acted upon on each trellis state, in order to determine the best previous state and, therefore, the updated history path. Furthermore, the quantization accumulated cost is updated for each state; in particular, this coatis consistent with the metric use for each LSP difference quantization.</li>
</ul></p>
<p id="p0105" num="0105">Note that the TCVQ generalization for the LSP multi-dimensional predictor case and for the reflection coefficients case can be derived in a straightforward manner following the corresponding TCQ descriptions.</p>
<p id="p0106" num="0106">Finally, also the trellis level reoptimization procedure can be carried out in an analogous way as for the TCQ case. In particular, the vector clusters can be constructed in an iterative way, as<!-- EPO <DP n="28"> --> function of the different trellis states and of the corresponding encoding paths. These clusters are obtained as 'centroid' (according to a predetermined metric) of corresponding partitions of the input vector set.</p>
</description><!-- EPO <DP n="29"> -->
<claims id="claims01" lang="en">
<claim id="c-en-01-0001" num="0001">
<claim-text>Method for speech coding, comprising the steps of:
<claim-text>- receiving in input a set of LPC filter coefficients;</claim-text>
<claim-text>- quantizing said LPC filter coefficients set by means of vector quantization;<br/>
characterized in that it further comprises the steps of:
<claim-text>generating an expanded set of quantization levels;</claim-text>
<claim-text>pruning said set of quantization levels using a TCQ technique.</claim-text></claim-text><!-- EPO <DP n="30"> --></claim-text></claim>
<claim id="c-en-01-0002" num="0002">
<claim-text>Method according to claim 1 characterized by a variable bit allocation at each quantization step.</claim-text></claim>
<claim id="c-en-01-0003" num="0003">
<claim-text>Method according to claim 2 characterized by the fact that a bit rate increase is obtained by adding one or more parallel transitions to the state branches, given a certain trellis topology.</claim-text></claim>
<claim id="c-en-01-0004" num="0004">
<claim-text>Method according to claim 2 characterized by the fact that a bit rate decrease is obtained by deleting one or more state branches, given a certain trellis topology.</claim-text></claim>
<claim id="c-en-01-0005" num="0005">
<claim-text>Method according to claim 1 characterized by the fact that of each quantization step the quantization error accumulated in quantizing the previous steps can be monitored and, eventually compensated.</claim-text></claim>
<claim id="c-en-01-0006" num="0006">
<claim-text>Method according to claim 1 characterized by said LPC filter coefficients being the LSP parameters.</claim-text></claim>
<claim id="c-en-01-0007" num="0007">
<claim-text>Method according to claims 5 and 6 characterized by the fact that at each quantization step each trellis state is assigned a history path, each path branch corresponding to a pointer to the quantization level of the corresponding LSP value.</claim-text></claim>
<claim id="c-en-01-0008" num="0008">
<claim-text>Method according to claim 7 characterized by the fact the<!-- EPO <DP n="31"> --> history path contains the information associated to each quantization step.</claim-text></claim>
<claim id="c-en-01-0009" num="0009">
<claim-text>Method according to claim 6 characterized by the fact that intraframe correlation can be exploited in quantizing the LSP parameters, by means of one dimensional differential prediction schemes along the frequency direction.</claim-text></claim>
<claim id="c-en-01-0010" num="0010">
<claim-text>Method according to claim 6 characterized by the fact that interframe correlation can be exploited in quantizing the LSP parameters, by means of one dimensional differential prediction schemes along the time direction.</claim-text></claim>
<claim id="c-en-01-0011" num="0011">
<claim-text>Method according to claim 6 characterized by the fact that both interframe correlation and intraframe correlation can be exploited in quantizing the LSP parameters, by means of multi dimensional differential prediction schemes.</claim-text></claim>
<claim id="c-en-01-0012" num="0012">
<claim-text>Method according to claim 1 characterized by said LPC filter coefficients being the RC parameters.</claim-text></claim>
<claim id="c-en-01-0013" num="0013">
<claim-text>Method according to claims 12 and 5 characterized by comprising the steps of defining trellis topology, defining quantization level number at each quantization step, and computing each RC as function of each particular state.</claim-text></claim>
<claim id="c-en-01-0014" num="0014">
<claim-text>Method according to claim 1 characterized by the fact that the quantization error is computed using a metric that has the property of being additive at each quantization step.</claim-text></claim>
<claim id="c-en-01-0015" num="0015">
<claim-text>Method according to claim 14 characterized by comprising the steps of reconstructing the encoding path in<!-- EPO <DP n="32"> --> correspondence of each trellis state, obtaining a set of quantized LPC parameter vectors, obtaining for each vector, a corresponding representation in terms of LPC cepstral coefficients, measuring the cepstral distance with respect to the cepstral coefficient representation of the unquantized model, choosing the trellis parameters that define the LPC vector with optimal cepstral distance.<!-- EPO <DP n="33"> --></claim-text></claim>
<claim id="c-en-01-0016" num="0016">
<claim-text>Method according to claim 1, characterized by adopting a TCQ training procedure for a quantization level reoptimization.</claim-text></claim>
<claim id="c-en-01-0017" num="0017">
<claim-text>Method according to claim 16, characterized by further comprising the step of starting from a set of quantization values for each state subset of the trellis, adopting an iterative procedure with a training sequence of LSP as input, assigning to each input LSP vector a partition corresponding to the obtained TCQ path, taking a mean of each partition value, assigning said mean to the corresponding path state branch.</claim-text></claim>
<claim id="c-en-01-0018" num="0018">
<claim-text>Method for speech coding, comprising the steps of:
<claim-text>- receiving in input a set of LPC filter coefficients;</claim-text>
<claim-text>- quantizing said LPC filter coefficients set by means of vector quantization;<br/>
characterized in that it further comprises the steps of:</claim-text>
<claim-text>- generating an expanded set of quantization levels;</claim-text>
<claim-text>- pruning said set of quantization levels using a TCVQ technique.</claim-text></claim-text></claim>
<claim id="c-en-01-0019" num="0019">
<claim-text>Speech coder based on LPC techniques, comprising means for quantizing an LPC filter coefficients set by means of vector quantization, and<br/>
further comprising:
<claim-text>- means for generating an expanded set of quantization levels;</claim-text>
<claim-text>- means for pruning said set of quantization levels using a TCQ technique.</claim-text></claim-text></claim>
</claims><!-- EPO <DP n="34"> -->
<claims id="claims02" lang="de">
<claim id="c-de-01-0001" num="0001">
<claim-text>Verfahren zur Sprachcodierung, das die Schritte umfasst:
<claim-text>- Empfangen einer Menge von LPC-Filterkoeffizienten am Eingang;</claim-text>
<claim-text>- Quantisieren der LPC-Filterkoeffizientenmenge mittels Vektorquantisierung;<br/>
dadurch gekennzeichnet, dass es ferner die Schritte umfasst
<claim-text>Erzeugen einer erweiterten Menge von Quantisierungsebenen; Beschneiden der Menge von Quantisierungsebenen unter Verwendung einer TCQ-Technik.</claim-text></claim-text></claim-text></claim>
<claim id="c-de-01-0002" num="0002">
<claim-text>Verfahren nach Anspruch 1, gekennzeichnet durch eine variable Bitzuordnung bei jedem Quantisierungsschritt.</claim-text></claim>
<claim id="c-de-01-0003" num="0003">
<claim-text>Verfahren nach Anspruch 2, dadurch gekennzeichnet, dass eine Bitratenerhöhung durch Hinzufügen eines oder mehrerer paralleler Übergänge in den Zustandszweigen erhalten wird, wenn eine bestimmte Trellis-Topologie gegeben ist.</claim-text></claim>
<claim id="c-de-01-0004" num="0004">
<claim-text>Verfahren nach Anspruch 2, dadurch gekennzeichnet, dass eine Bitratenverringerung erhalten wird, indem einer oder mehrere Zustandszweige gelöscht werden, wenn eine bestimmte Trellis-Topologie gegeben ist.<!-- EPO <DP n="35"> --></claim-text></claim>
<claim id="c-de-01-0005" num="0005">
<claim-text>Verfahren nach Anspruch 1, dadurch gekennzeichnet, dass von jedem Quantisierungsschritt der beim Quantisieren der vorherigen Schritte aufgelaufene Quantisierungsfehler überwacht und schließlich kompensiert werden kann.</claim-text></claim>
<claim id="c-de-01-0006" num="0006">
<claim-text>Verfahren nach Anspruch 1, gekennzeichnet durch die LPC-Filterkoeffizienten, die die LSP-Parameter sind.</claim-text></claim>
<claim id="c-de-01-0007" num="0007">
<claim-text>Verfahren nach den Ansprüchen 5 und 6, dadurch gekennzeichnet, dass bei jedem Quantisierungsschritt jedem Trellis-Zustand ein Pfad der geschichtlichen Entwicklung zugewiesen wird, wobei jeder Pfadzweig einem Zeiger zur Quantisierungsebene des entsprechenden LSP-Wertes entspricht.</claim-text></claim>
<claim id="c-de-01-0008" num="0008">
<claim-text>Verfahren nach Anspruch 7, dadurch gekennzeichnet, dass der Pfad der geschichtlichen Entwicklung die zu jedem Quantisierungsschritt gehörige Information enthält.</claim-text></claim>
<claim id="c-de-01-0009" num="0009">
<claim-text>Verfahren nach Anspruch 6, dadurch gekennzeichnet, dass die rahmeninterne Korrelation beim Quantisieren der LSP-Parameter mittels eindimensionaler differentieller Vorhersagemethoden entlang der Frequenzrichtung ausgenutzt werden kann.</claim-text></claim>
<claim id="c-de-01-0010" num="0010">
<claim-text>Verfahren nach Anspruch 6, dadurch gekennzeichnet, dass die Zwischenrahmenkorrelation beim Quantisieren der LSP-Parameter mittels eindimensionaler differentieller Vorhersagemethoden entlang der Zeitrichtung ausgenutzt werden kann.<!-- EPO <DP n="36"> --></claim-text></claim>
<claim id="c-de-01-0011" num="0011">
<claim-text>Verfahren nach Anspruch 6, dadurch gekennzeichnet, dass sowohl die rahmeninterne Korrelation als auch die Zwischenrahmenkorrelation beim Quantisieren der LSP-Parameter mittels mehrdimensionaler differentieller Vorhersagemethoden ausgenutzt werden kann.</claim-text></claim>
<claim id="c-de-01-0012" num="0012">
<claim-text>Verfahren nach Anspruch 1, gekennzeichnet durch die LPC-Filterkoeffizienten, die die RC-Parameter sind.</claim-text></claim>
<claim id="c-de-01-0013" num="0013">
<claim-text>Verfahren nach den Ansprüche 12 und 5, dadurch gekennzeichnet, dass es die Schritte des Definierens einer Trellis-Topologie, des Definierens einer Anzahl von Quantisierungsebenen bei jedem Quantisierungsschritt und des Berechnens jedes RC als Funktion jedes besonderen Zustands umfasst.</claim-text></claim>
<claim id="c-de-01-0014" num="0014">
<claim-text>Verfahren nach Anspruch 1, dadurch gekennzeichnet, dass der Quantisierungsfehler unter Verwendung einer Metrik berechnet wird, die die Eigenschaft hat, dass sie bei jedem Quantisierungsschritt additiv ist.</claim-text></claim>
<claim id="c-de-01-0015" num="0015">
<claim-text>Verfahren nach Anspruch 14, dadurch gekennzeichnet, dass es die Schritte des Rekonstruierens des Codierpfades in Übereinstimmung mit jedem Trellis-Zustand, des Erhaltens eines Satzes von quantisierten LPC-Parametervektoren, des Erhaltens einer entsprechenden Darstellung in Form von LPC-Cepstrumkoeffizienten für jeden Vektor, des Messens des cepstralen Abstandes hinsichtlich der Cepstrum-Koeffizientendarstellung des nicht quantisierten Modells, des Auswählens der Trellis-Parameter, die den LPC-Vektor mit dem optimalen cepstralen Abstand definieren, umfasst.<!-- EPO <DP n="37"> --></claim-text></claim>
<claim id="c-de-01-0016" num="0016">
<claim-text>Verfahren nach Anspruch 1, gekennzeichnet durch die Übernahme einer TCQ-Schulungsprozedur für eine erneute Optimierung der Quantisierungsebenen.</claim-text></claim>
<claim id="c-de-01-0017" num="0017">
<claim-text>Verfahren nach Anspruch 16, dadurch gekennzeichnet, dass es ferner die Schritte umfasst: Ausgehen von einer Menge von Quantisierungswerten für jede Zustandsuntermenge des Trellis, Übernahme einer iterativen Prozedur mit einer Schulungsfolge von LSP als Eingabe, Zuweisung einer dem erhaltenen TCQ-Pfad entsprechenden Verteilung zu jedem eingegebenen LSP-Vektor, Nehmen eines Mittelwerts jedes Verteilungswertes,Zuweisen des Mittelwertes zu dem entsprechenden Zustandszweig des Pfades.</claim-text></claim>
<claim id="c-de-01-0018" num="0018">
<claim-text>Verfahren zur Sprachcodierung, das die Schritte umfasst:
<claim-text>- Empfangen einer Menge von LPC-Filterkoeffizienten am Eingang;</claim-text>
<claim-text>- Quantisieren der Menge der LPC-Filterkoeffizienten mittels Vektorquantisierung,<br/>
dadurch gekennzeichnet, dass es ferner die Schritte umfasst:</claim-text>
<claim-text>- Erzeugen einer erweiterten Menge von Quantisierungsebenen,</claim-text>
<claim-text>- Beschneiden der Menge von Quantisierungsebenen unter Verwendung einer TCVQ-Technik.</claim-text></claim-text></claim>
<claim id="c-de-01-0019" num="0019">
<claim-text>Sprachcodierer auf der Basis von LPC-Techniken, der Mittel zum Quantisieren einer Menge von LPC-Filterkoeffizienten mittels Vektorquantisierung umfasst und ferner umfasst:
<claim-text>- Mittel zum Erzeugen einer erweiterten Menge von<!-- EPO <DP n="38"> --> Quantisierungsebenen;</claim-text>
<claim-text>- Mittel zum Beschneiden der Menge von Quantisierungsebenen unter Verwendung einer TCQ-Technik.</claim-text></claim-text></claim>
</claims><!-- EPO <DP n="39"> -->
<claims id="claims03" lang="fr">
<claim id="c-fr-01-0001" num="0001">
<claim-text>Procédé de codage de la parole, comprenant les étapes consistant à :
<claim-text>- recevoir en entrée un ensemble de coefficients de filtre LPC;</claim-text>
<claim-text>- quantifier ledit ensemble de coefficients de filtre LPC au moyen d'une quantification vectorielle;<br/>
caractérisé en ce qu'il comprend en outre les étapes consistant à :</claim-text>
<claim-text>- générer un ensemble étendu de niveaux de quantification;</claim-text>
<claim-text>- réduire ledit ensemble de niveaux de quantification à l'aide d'une technique de TCQ.</claim-text></claim-text></claim>
<claim id="c-fr-01-0002" num="0002">
<claim-text>Procédé selon la revendication 1 caractérisé par une allocation binaire variable à chaque étape de quantification.</claim-text></claim>
<claim id="c-fr-01-0003" num="0003">
<claim-text>Procédé selon la revendication 2, caractérisé en ce qu'une augmentation du débit binaire est obtenue en ajoutant une ou plusieurs transitions parallèles aux branches d'état, étant donné une certaine topologie de treillis.</claim-text></claim>
<claim id="c-fr-01-0004" num="0004">
<claim-text>Procédé selon la revendication 2, caractérisé en ce qu'une diminution du débit binaire est obtenue en supprimant une ou plusieurs branches d'état, étant donné une certaine topologie de treillis.</claim-text></claim>
<claim id="c-fr-01-0005" num="0005">
<claim-text>Procédé selon la revendication 1, caractérisé en ce que de chaque étape de quantification, l'erreur de quantification cumulée pendant la quantification des étapes précédentes peut être surveillée et, enfin compensée.<!-- EPO <DP n="40"> --></claim-text></claim>
<claim id="c-fr-01-0006" num="0006">
<claim-text>Procédé selon la revendication 1, caractérisé en ce que lesdits coefficients de filtre LPC sont les paramètres LSP.</claim-text></claim>
<claim id="c-fr-01-0007" num="0007">
<claim-text>Procédé selon les revendications 5 et 6, caractérisé en ce qu'à chaque étape de quantification, à chaque état de treillis est affecté un chemin d'historique, chaque branche de chemin correspondant à un pointeur vers un niveau de quantification de la valeur LSP correspondante.</claim-text></claim>
<claim id="c-fr-01-0008" num="0008">
<claim-text>Procédé selon la revendication 7 caractérisé en ce que le chemin d'historique contient les informations associées à chaque étape de quantification.</claim-text></claim>
<claim id="c-fr-01-0009" num="0009">
<claim-text>Procédé selon la revendication 6, caractérisé en ce que la corrélation intra-trame peut être exploitée en quantifiant les paramètres LSP, au moyen d'une méthode de prédiction différentielle dimensionnelle suivant la direction de fréquence.</claim-text></claim>
<claim id="c-fr-01-0010" num="0010">
<claim-text>Procédé selon la revendication 6, caractérisé en ce que la corrélation inter-trame peut être exploitée en quantifiant les paramètres LSP, au moyen d'une méthode de prédiction différentielle dimensionnelle suivant la direction de temps.</claim-text></claim>
<claim id="c-fr-01-0011" num="0011">
<claim-text>Procédé selon la revendication 6, caractérisé en ce que tant la corrélation inter-trame que la corrélation intra-trame peuvent être exploitées en quantifiant les paramètres LSP, au moyen de méthodes de prédiction différentielle multidimensionnelle.</claim-text></claim>
<claim id="c-fr-01-0012" num="0012">
<claim-text>Procédé selon la revendication 1, caractérisé en ce que lesdits coefficients de filtre LPC sont les paramètres RC.</claim-text></claim>
<claim id="c-fr-01-0013" num="0013">
<claim-text>Procédé selon les revendications 12 et 5, caractérisé en ce qu'il comprend les étapes consistant à définir la topologie de treillis, définir le nombre de niveaux de<!-- EPO <DP n="41"> --> quantification à chaque étape de quantification, et calculer chaque RC en fonction de chaque état particulier.</claim-text></claim>
<claim id="c-fr-01-0014" num="0014">
<claim-text>Procédé selon la revendication 1, caractérisé en ce que l'erreur de quantification est calculée à l'aide d'une métrique qui présente la propriété d'être additive à chaque étape de quantification.</claim-text></claim>
<claim id="c-fr-01-0015" num="0015">
<claim-text>Procédé selon la revendication 14, caractérisé en ce qu'il comprend les étapes consistant à reconstruire le chemin de codage en correspondance avec chaque état de treillis, obtenir un ensemble de vecteurs de paramètres LPC quantifiés, obtenir pour chaque vecteur, une représentation correspondante en fonction des coefficients cepstraux LPC, mesurer la distance cepstrale par rapport à la représentation des coefficients cepstraux du modèle non quantifié, choisir les paramètres de treillis qui définissent le vecteur LPC avec la distance cepstrale optimale.</claim-text></claim>
<claim id="c-fr-01-0016" num="0016">
<claim-text>Procédé selon la revendication 1, caractérisé en ce qu'il adopte une procédure d'entraînement TCQ pour une réoptimisation des niveaux de quantification.</claim-text></claim>
<claim id="c-fr-01-0017" num="0017">
<claim-text>Procédé selon la revendication 16, caractérisé en ce qu'il comprend en outre l'étape consistant à démarrer à partir d'un ensemble de valeurs de quantification pour chaque sous-ensemble d'états du treillis, adopter une procédure itérative avec une séquence d'entraînement de la LSP en entrée, affecter à chaque vecteur LSP d'entrée une partition correspondant au chemin TCQ obtenu, prendre une moyenne de chaque valeur de partition, affecter ladite moyenne à la branche d'état de chemin correspondante.<!-- EPO <DP n="42"> --></claim-text></claim>
<claim id="c-fr-01-0018" num="0018">
<claim-text>Procédé de codage de la parole, comprenant les étapes consistant à :
<claim-text>- recevoir en entrée un ensemble de coefficients de filtre LPC;</claim-text>
<claim-text>- quantifier ledit ensemble de coefficients de filtre LPC au moyen d'une quantification vectorielle;<br/>
caractérisé en ce qu'il comprend en outre les étapes consistant à :</claim-text>
<claim-text>- générer un ensemble étendu de niveaux de quantification;</claim-text>
<claim-text>- réduire ledit ensemble de niveaux de quantification à l'aide d'une technique de TCVQ.</claim-text></claim-text></claim>
<claim id="c-fr-01-0019" num="0019">
<claim-text>Codeur de parole basé sur des techniques LPC, comprenant des moyens pour quantifier un ensemble de coefficients de filtre LPC au moyen d'une quantification vectorielle, et comprenant en outre :
<claim-text>- des moyens pour générer un ensemble étendu de niveaux de quantification;</claim-text>
<claim-text>- des moyens pour réduire ledit ensemble de niveaux de quantification à l'aide d'une technique de TCQ.</claim-text></claim-text></claim>
</claims><!-- EPO <DP n="43"> -->
<drawings id="draw" lang="en">
<figure id="f0001" num=""><img id="if0001" file="imgf0001.tif" wi="146" he="245" img-content="drawing" img-format="tif"/></figure><!-- EPO <DP n="44"> -->
<figure id="f0002" num=""><img id="if0002" file="imgf0002.tif" wi="143" he="241" img-content="drawing" img-format="tif"/></figure>
</drawings>
</ep-patent-document>
