Perceptual weighting device and method for efficient coding...

Data processing: speech signal processing – linguistics – language – Speech signal processing – Psychoacoustic

Reexamination Certificate

Rate now

[ 0.00 ] – not rated yet Voters 0 Comments 0

Details Perceptual weighting device and method for efficient coding... Perceptual weighting device and method for efficient coding...

: 2001-06-20
: 2004-10-19
: {haeck over (S)}mits, T{overscore (a)}livaldis Ivars (Department: 2655)
: Data processing: speech signal processing, linguistics, language
: Speech signal processing
: Psychoacoustic

: C704S219000, C704S262000, C704S224000
: Reexamination Certificate
: active
: 06807524
: ABSTRACT:

BACKGROUND OF THE INVENTION
1. Field of the invention
The present invention relates to a perceptual weighting device and method for producing a perceptually weighted signal in response to a wideband signal (0-7000 Hz) in order to reduce a difference between a weighted wideband signal and a subsequently synthesized weighted wideband signal.
2. Brief description of the prior art
The demand for efficient digital wideband speech/audio encoding techniques with a good subjective quality/bit rate trade-off is increasing for numerous applications such as audio/video teleconferencing, multimedia, and wireless applications, as well as Internet and packet network applications. Until recently, telephone bandwidths filtered in the range 200-3400 Hz were mainly used in speech coding applications. However, there is an increasing demand for wideband speech applications in order to increase the intelligibility and naturalness of the speech signals. A bandwidth in the range 50-7000 Hz was found sufficient for delivering a face-to-face speech quality. For audio signals, this range gives an acceptable audio quality, but is still lower than the CD quality which operates on the range 20-20000 Hz.
A speech encoder converts a speech signal into a digital bitstream which is transmitted over a communication channel (or stored in a storage medium). The speech signal is digitized (sampled and quantized with usually 16-bits per sample) and the speech encoder has the role of representing these digital samples with a smaller number of bits while maintaining a good subjective speech quality. The speech decoder or synthesizer operates on the transmitted or stored bit stream and converts it back to a sound signal.
One of the best prior art techniques capable of achieving a good quality/bit rate trade-off is the so-called Code Excited Linear Prediction (CELP) technique. According to this technique, the sampled speech signal is processed in successive blocks of L samples usually called frames where L is some predetermined number (corresponding to 10-30 ms of speech). In CELP, a linear prediction (LP) synthesis filter is computed and transmitted every frame. The L-sample frame is then divided into smaller blocks called subframes of size N samples, where L=kN and k is the number of subframes in a frame (N usually corresponds to 4-10 ms of speech). An excitation signal is determined in each subframe, which usually consists of two components: one from the past excitation (also called pitch contribution or adaptive codebook) and the other from an innovative codebook (also called fixed codebook). This excitation signal is transmitted and used at the decoder as the input of the LP synthesis filter in order to obtain the synthesized speech.
An innovative codebook in the CELP context, is an indexed set of N-sample-long sequences which will be referred to as N-dimensional codevectors. Each codebook sequence is indexed by an integer k ranging from 1 to M where M represents the size of the codebook often expressed as a number of bits b, where M=2
b
.
To synthesize speech according to the CELP technique, each block of N samples is synthesized by filtering an appropriate codevector from a codebook through time varying filters modelling the spectral characteristics of the speech signal. At the encoder end, the synthesis output is computed for all, or a subset, of the codevectors from the codebook (codebook search). The retained codevector is the one producing the synthesis output closest to the original speech signal according to a perceptually weighted distortion measure. This perceptual weighting is performed using a so-called perceptual weighting filter, which is usually derived from the LP synthesis filter.
The CELP model has been very successful in encoding telephone band sound signals, and several CELP-based standards exist in a wide range of applications, especially in digital cellular applications. In the telephone band, the sound signal is band-limited to 200-3400 Hz and sampled at 8000 samples/sec. In wideband speech/audio applications, the sound signal is band-limited to 50-7000 Hz and sampled at 16000 samples/sec.
Some difficulties arise when applying the telephone-band optimized CELP model to wideband signals, and additional features need to be added to the model in order to obtain high quality wideband signals. Wideband signals exhibit a much wider dynamic range compared to telephone-band signals, which results in precision problems when a fixed-point implementation of the algorithm is required (which is essential in wireless applications). Furthermore, the CELP model will often spend most of its encoding bits on the low-frequency region, which usually has higher energy contents, resulting in a low-pass output signal. To overcome this problem, the perceptual weighting filter has to be modified in order to suit wideband signals, and pre-emphasis techniques which boost the high frequency regions become important to reduce the dynamic range, yielding a simpler fixed-point implementation, and to ensure a better encoding of the higher frequency contents of the signal.
In CELP-type encoders, the optimum pitch and innovative parameters are searched by minimizing the mean squared error between the input speech and synthesized speech in a perceptually weighted domain. This is equivalent to minimizing the error between the weighted input speech and weighted synthesis speech, where the weighting is performed using a filter having a transfer function W(z) of the form:
W
(
z
)=
A
(
z/g
1
)
/A
(
z/g
2
) where 0<&Ggr;
2
<&Ggr;
1≦
1.
In analysis-by-synthesis (AbS) coders, analysis show that the quantization error is weighted by the inverse of the weighting filter, W
−1
(z), which exhibits some of the formant structure in the input signal. Thus, the masking property of the human ear is exploited by shaping the error, so that it has more energy in the formant regions, where it will be masked by the strong signal energy present in those regions. The amount of weighting is controlled by the factors &Ggr;
1
and &Ggr;
2
.
This filter works well with telephone band signals. However, it was found that this filter is not suitable for efficient perceptual weighting when it was applied to wideband signals. It was found that this filter has inherent limitations in modelling the formant structure and the required spectral tilt concurrently. The spectral tilt is more pronounced in wideband signals due to the wide dynamic range between low and high frequencies. It was suggested to add a tilt filter into filter W(z) in order to control the tilt and formant weighting separately.
OBJECT OF THE INVENTION
An object of the present invention is therefore to provide a perceptual weighting device and method adapted to wideband signals, using a modified perceptual weighting filter to obtain a high quality reconstructed signal, these device and method enabling fixed point algorithmic implementation.
SUMMARY OF THE INVENTION
More specifically, in accordance with the present invention, there is provided a perceptual weighting device for producing a perceptually weighted signal in response to a wideband signal in order to reduce a difference between a weighted wideband signal and a subsequently synthesized weighted wideband signal. This perceptual weighting device comprises:
a) a signal preemphasis filter responsive to the wideband signal for enhancing the high frequency content of the wideband signal to thereby produce a preemphasised signal;
b) a synthesis filter calculator responsive to the preemphasised signal for producing synthesis filter coefficients; and
c) a perceptual weighting filter, responsive to the preemphasised signal and the synthesis filter coefficients, for filtering the preemphasised signal in relation to the synthesis filter coefficients to thereby produce the perceptually weighted signal. The perceptual weighting filter has a transfer function with fixed denominator whereby weighting of the wideband signal in a formant region is substantially decoupled from a spectral tilt of that wideband signal.

Affiliated with

Bessette Bruno

Inventor

[ 0.00 ] – not rated yet Voters 0 Comments 0

Lefebvre Roch

Inventor

[ 0.00 ] – not rated yet Voters 0 Comments 0

Salami Redwan

Inventor

[ 0.00 ] – not rated yet Voters 0 Comments 0

Also associated with

Voiceage Corporation

Corporate Assignee

[ 0.00 ] – not rated yet Voters 0 Comments 0

Wozniak James S.

Examiner

[ 0.00 ] – not rated yet Voters 0 Comments 0

{haeck over (S)}mits T{overscore (a)}livaldis Ivars

Examiner

[ 0.00 ] – not rated yet Voters 0 Comments 0

LandOfFree

Say what you really think

Search LandOfFree.com for the USA inventors and patents. Rate them and share your experience with other people.

Rating

Perceptual weighting device and method for efficient coding... does not yet have a rating. At this time, there are no reviews or comments for this patent.
If you have personal experience with Perceptual weighting device and method for efficient coding..., we encourage you to share that experience with our LandOfFree.com community. Your opinion is very important and Perceptual weighting device and method for efficient coding... will most certainly appreciate the feedback.

Rate now

Comments { 0 }

Profile ID: LFUS-PAI-O-3321929

All data on this website is collected from public sources. Our data reflects the most accurate information available at the time of publication.

Canada

Charities
Companies
MP Candidates
Patents
Employee Salary Disclosure

World

Places of the World
Scientific Papers

United States

Banks
Companies
Counties
Patents
Employee Salary Disclosure