Signal Processing Techniques in Wireless Systems
Signal Processing Techniques in Wireless Systems
Wireless Systems 14
14.1 Introduction
ignal processing techniques such as diversity, equalization, speech, and channel coding
14.2 Diversity
Diversity refers to a technique for improving the transmission of a signal, by receiving and
processing multiple versions of the same transmitted signal. The multiple received versions can
be the result of signals following different propagation paths (spatial diversity), being transmitted
at different times (time diversity) or frequencies (frequency diversity).
332 Mobile Cellular Communication
Diversity is a commonly used technique in mobile radio systems to combat signal fading. If
several replicas of the same information carrying signal are received over multiple channels with
comparable strengths which exhibit independent fading, then there is a good likelihood that at
least one or more of these received signals will not be faded at any given instant. This makes it
possible to deliver adequate signal level to the receiver. Without diversity techniques, in noise-
limited conditions, the transmitter would have to deliver a much higher power level to protect
the link during the short intervals when the channel is severely faded. In mobile radio, the power
available on the reverse link is severely limited by the battery capacity of hand-held subscriber
units. Diversity methods play a crucial role in reducing transmitted power needs. Also, cellular
communication networks are mostly interference limited and mitigation of channel fading
through use of diversity can translate into reduced variability of carrier to interference ratio (C/I),
which in turn means lower C/I margin and hence better reuse factors and higher system capacity.
E ⎣⎡[ r1 − r1 ][ r2 − r2 ]⎦⎤
r = (14.1)
E [ r1 − r1 ] E [ r2 − r2 ]
2 2
where E is the expectance and r1 and r2 represent the instantaneous envelope levels of the
normalized signals at the two receivers, and r1 and r2 are their respective means.
It has been shown that a cross-correlation of 0.7 between signal envelopes is sufficient to
provide a reasonable degree of diversity gain. Depending on the type of diversity employed, these
diversity channels must be sufficiently separated along the appropriate diversity dimension. For
Signal Processing in Wireless Systems 333
–90 3
–95
–100
Signal level in dB
–105
–110 2 r2(t)
–115 1
r1(t)
–120
–125
–130
Time in milliseconds
Figure 14.1 Example of diversity combining. The two independently fading signals are 1 and 2.
Signal 3 is the result of selecting the strongest signal
spatial diversity, the antennas should be separated by more than the coherence distance to ensure
a cross-correlation of less than 0.7.
Space diversity
Space diversity, also known as antenna diversity, is one of the most popular forms of receiver
diversity schemes widely used in wireless systems. It is easy to implement and does not require
additional frequency spectrum resources. Space diversity is exploited on the reverse link at
the base station receiver by spacing antennas apart so as to obtain sufficient decorrelation.
334 Mobile Cellular Communication
TX and RX TX and RX
0.7l
RX only
RX only
The key for obtaining minimum uncorrelated fading of antenna outputs is adequate spacing
of the antennas. The required spacing depends on the degree of multipath angle spread. For
example, if the multipath signals arrive from all directions in the azimuth, as is usually the
case at the mobile, antenna spacing (coherence distance) of the order of 0.5l to 0.8l is quite
adequate.
On the other hand, if the multipath angle spread is small, as in the case of base stations,
the coherence distance is much larger. Also, empirical measurements show a strong coupling
between antenna height and spatial correlation. Larger antenna heights imply larger coherence
distances. Typically, 10l–20l separation is adequate to achieve r = 0.7 at base stations in suburban
settings when the signals arrive from the broadside direction. The coherence distance can be
3–4 times larger for end fire arrivals. The end-fire problem is averted in base stations with tri-
sectored antennas as each sector needs to handle only signals arriving ±60° off the broadside.
The coherence distance depends strongly on the terrain. The space diversity topology is shown
in Figure 14.2. There is a statistical relationship between the spatial separation of two receiving
antennas and the likelihood of both receiving faded signals simultaneously.
Polarization diversity
In polarization diversity, horizontally and vertically polarized signals or left and right circular polarized
waves are transmitted simultaneously. Since these are uncorrelated in the mobile radio path, one of
these signals will provide strong received level after fading. In mobile radio environments, signals
transmitted on orthogonal polarizations exhibit low fade correlation and, therefore, offer potential
for diversity combining. Polarization diversity can be obtained in two ways:
Frequency diversity
Another technique to obtain decorrelated diversity branches is to transmit the same signal
over different frequencies. The frequency separation between carriers should be larger than the
coherence bandwidth. The coherence bandwidth, of course, depends on the multipath delay
spread of the channel. The larger the delay spread, the smaller the coherence bandwidth and
more closely we can space the frequency diversity channels. Clearly, frequency diversity is
an explicit diversity technique and needs additional frequency spectrum. A common form of
frequency diversity is multicarrier (also known as multitone) modulation.
Signal Processing in Wireless Systems 335
This technique involves sending redundant data over a number of closely spaced carriers to
benefit from frequency diversity, which is then exploited by applying interleaving and channel
coding/forward error correction (FEC) across the carriers. Another technique is to use frequency
hopping wherein the interleaved and channel coded data stream is transmitted with widely
separated frequencies from burst to burst. The wide frequency separation is chosen to guarantee
independent fading from burst to burst. Some examples of systems employing frequency diversity
include spread spectrum systems, such as frequency-hopping spread spectrum (FHSS) and multi-
carrier spread spectrum (MCSS) systems.
Time diversity
In time diversity, the desired message is transmitted repeatedly over several time periods having
separation between adjacent time slots larger than the channel coherence time. In mobile
communications channels, the motion of the mobile together with scattering in the vicinity of
the mobile causes time selective fading of the signal with Rayleigh fading statistics for the signal
envelope. Signal fade levels separated by the coherence time show low correlation and can be
used as diversity branches, if the same signal can be transmitted at multiple instants separated
by the coherence time.
The coherence time depends on the Doppler spread of the signal, which in turn is a function
of the mobile speed and the carrier frequency. Time diversity is usually exploited via interleaving,
FEC coding, and automatic request for repeat (ARQ). These are sophisticated techniques to exploit
channel coding and time diversity. One fundamental drawback with time diversity approaches
is the delay needed to collect the repeated or interleaved transmissions. If the coherence time is
large, as for example, when the vehicle is slow moving, the required delay becomes too large to
be acceptable for interactive voice conversation.
combining and requires an infrastructure that mediates the signals from the local antennas or
receivers to a central receiver or decoder.
Transmitter macro diversity may be a form of simultaneous broadcasting, where the same
signal is sent from several nodes. If the signals are sent over the same physical channel (e.g., the
channel frequency and spreading sequence), the transmitters are said to form a single frequency
network, a term used especially in the broadcasting world. The aim is to combat fading and to
increase the received signal strength and signal quality in exposed positions in between the
base stations or access points. Macro diversity may also facilitate efficient broadcasting and
multicasting services, where the same frequency channel can be used for all transmitters sending
the same information. The diversity scheme may be based on transmitter (downlink) macro
diversity and/or receiver (uplink) macro diversity.
Selection combining
Selection combining is the simplest and perhaps the most frequently used form of diversity
combining. With selective diversity, one best signal is chosen based on the received signal
strengths from the set of diversity branches. The receiver monitors the SNR value of each diversity
branch and selects the one with the maximum SNR value for signal detection. It is much easier to
implement without much performance degradation. It is generally employed for the reverse link
transmission where the diversity branches can be physically located in different base stations.
Figure 14.3 depicts a simple block diagram of the selective diversity combining technique.
Consider M number of independent fading signals received by multiple receiver antennas.
There are M-branch receivers comprising of coherent demodulators. The output of demodulators
is presented to a logic circuit, which selects the particular branch receiver output having the largest
SNR value of the received signal. Conceptually, selection diversity combining is the simplest
Receiver 1
Receiver 2 Output
Logic
circuit
Receiver M
form of “space diversity on receive” technique. The selection combining procedure requires that
the receiver outputs are monitored continuously. At each instant of time, the receiver with the
largest instantaneous SNR value is selected. From a practical implementation point of view, such
a selective combining procedure is difficult.
There is an alternate method by adopting a scanning version of the selective combining
procedure. Firstly, the receiver with the strongest output signal is selected and maintained till its
instantaneous SNR value does not drop below a pre-defined threshold SNR value. Then, a new
receiver that offers the strongest output signal is selected and the selection procedure is repeated.
This method has an advantage that it requires only one receiver and is very simple to implement.
The performance of the selective combining technique is not optimum because it ignores the
information available from all the diversity branches. But there is exception for the particular
selected receiver branch that produces the instantaneous SNR value greater than the pre-defined
threshold SNR value. The probability, PM (Si), that all independent diversity branches receive
signals, which are simultaneously less than specified threshold SNR, Si, is given by
(
PM (Si ) = 1 − e si / sr )
M
(14.2)
where
Sr is the average received SNR value
M is the number of independent diversity branches
It can be seen that selective diversity can greatly improve the performance of the bit error rate
(BER). The performance improvement is more significant when M is increased from 1 to 2 than it
is increased from 2 to 4 and 4 to 8.
Receiver 1
Receiver 2
Output
Linear
combiner
Receiver M
bandwidth Bc, which indicates the amount of bandwidth that will fade in a correlated fashion at
any instant of time. To define this correlation, consider the delay spread of the channel, Td, which
is the maximum duration of the mobile communication channel. The time index is a measure
of the time of arrival relative to the first multipath component at time 0. Often, the “direct
path” arrives first and subsequent paths represent paths reflected at increasing distances from the
receiver. The coherence bandwidth is given by
1
Bc ≈ (14.3)
Td
The coherence bandwidth is approximated at typically path powers less than 5–10 per cent of
the total power. Considering a communication system with bandwidth Bw. If Bc > Bw, the channel
between the transmitter and receiver is called a flat fading channel, and if Bc < Bw the channel
is called a frequency selective channel. Flat fading channels are problematic for systems without
time diversity, because a deep fade can result in a received signal that is below the background
noise level, making communication unreliable. The worst types of channel conditions for many
communication systems are slowly changing flat fading channels. This is due to the length of
time the receiver cannot reliably demodulate the bits sent by the transmitter.
TD can improve the receiver performance in the presence of flat fading. It reduces the
impact of fading by offering multiple independent copies of the digitally modulated waveform
at the receiver, where the chance that all copies are simultaneously in a fade is very small. TD
in radio communication using signals that originate from two or more independent sources has
been modulated with identical information bearing signals. It may vary in their transmission
characteristics at any given instant. It can help overcome the effects of fading, outages, and
circuit failures. When using diversity transmission and reception, the amount of received signal
improvement depends on the independence of the fading characteristics of the signal as well as
circuit outages and failures.
14.3 Equalization
Equalization means equalizing all parts of the received signal with respect to time delay and
frequency to the same attenuation level. This is because of the fact that the various components
of the signals suffer different levels of attenuation or fading when travelling through the medium.
This helps mitigate the ISI, fading effect, and so on. An equalizer is usually implemented at the
baseband or at the IF section in a receiver. In a typical wireless system, the RF communication
occurs in a pass band of bandwidth B with a centre frequency fc. Most of the processing such
as coding/decoding, modulation/demodulation, and synchronization are performed in the
baseband.
Equalization techniques which can combat and/or exploit the frequency selectivity of the
wireless channel are of enormous importance in the design of high data rate wireless systems.
The purpose of an equalizer is to reduce the ISI as much as possible to maximize the probability
of correct decisions.
Equalizers
Linear Non-linear
DFE MLSE
MAP
Transversal Lattice symbol
detector
Zero-forcing Gradient
LMS RLS
RLS Transversal Lattice
Fast RLS
Square root RLS LMS
Gradient Transversal channel
RLS
RLS estimator
Fast RLS
Square root RLS
LMS
RLS
Fast RLS
Square root RLS
represent the information the transmitter wants to send. Equalization techniques can be broadly
classified as linear and non-linear (see Fig. 14.5). These categories are determined from output of
an adaptive equalizer used for subsequent control of the equalizer. In general, the analogue signal
d(t) is processed by the decision-making device in the receiver. The decision maker determines
the value of the digital data bit being received and applies a slicing or thresholding operation (a
non-linear operation) in order to determine the value of d(t). If d(t) is not used in the feedback
path to adapt the equalizer, the equalization is linear. On the other hand, if d(t) is fed back
to change the subsequent outputs of the equalizer, the equalization is non-linear. Many filter
structures are used to implement linear and non-linear equalizers.
Equalization is used to overcome ISI due to channel time dispersion, and diversity is used to
overcome flat fading.
in near vicinity of the spectrum, so they are not very effective on channels having severe distortion.
The ISI can be completely removed, without taking into consideration the resulting noise enhacement.
Symbol spaced equalizers and fractionally spaced equalizers are examples of linear equalizers.
A symbol spaced linear equalizer (SSLE) consists of a tapped delay line that stores samples from
the input signal. Once per symbol period, the equalizer outputs a weighted sum of the values
in the delay line and updates the weights to prepare for the next symbol period. This class of
equalizer is called “symbol spaced” because the sample rates of the input and output are equal.
The algorithms for the weight setting and error calculation blocks are determined by the adaptive
algorithm chosen. The new set of weights depends on the current set of weights, the input
signal, the output signal, and for adaptive algorithms other than constant modulus algorithm, a
reference signal whose characteristics depend on the operation mode of the equalizer. In typical
applications, the equalizer begins in training mode to gather information about the channel, and
later switches to decision-directed mode.
A fractionally spaced equalizer (FSE) is a linear equalizer that is similar to a symbol spaced
linear equalizer. By contrast, however, a fractionally spaced equalizer receives Z input samples
before it produces one output sample and updates the weights, where Z is an integer. In many
applications, Z is 2. The output sample rate is 1/Tb, while the input sample rate is Z/Tb, where Tb
is the bit duration. The weight updating occurs at the output rate, which is slower. Figure 14.6
illustrates a general approach to implement an equalizer using a linear equalizer circuit. The
linear equalizer consists of a tapped delay line, which stores the input samples and then outputs
a weighted sum of the stored values once per symbol interval.
Some algorithms like Wiener algorithm are used to minimize the error, and the tap coefficients
are updated in preparation for the next symbol interval. If the time delay t is equal to symbol
interval TS (Symbol rate = RS), then the linear equalizer is called the symbol spaced equalizer. Such
equalizers are optimum if preceded by receiving filters matched to the operating characteristics
of the channels. However, in such practical cases, the channel characteristics are unknown and
consequently the equalizers are sub-optimum.
Equalized
output
Σ
Tap gain
adjustment
Feedback
filter
Sampled
received signal
+ −
Feed forward Decision
∑ device
filter Deleted
symbol output
∑
Error signal to
− + feedback filter
channel response need not be synthesized in the feed forward filter, therefore excessive noise
enhancement is avoided and sensitivity to sampler phase is decreased.
The advantage of a DFE implementation is the feedback filter, which is additionally working
to remove ISI, which operates on noiseless quantized levels and thus its output is free from
channel noise. A drawback of the DFE structure surfaces when an incorrect decision is applied
to the feedback filter. The DFE output reflects this error during the next few symbols as the
incorrect decision propagates through the feedback filter. Under this condition, there is a greater
likelihood of more incorrect decisions following the first one, producing a condition known as
error propagation. On most channels of interest, the error rate is so low enough that the overall
performance degradation is less significant.
Channel estimator
Sampled
received
signal
Delay Delay Delay
Detected
MLSE using viterbi algorithm symbols
The sampled impulse response of the channel, taken at symbol intervals, is measured with
the adaptive channel estimator. It is then compared with the sequence of the sampled received
signal with all possible received sequences and determines the most likely transmitted sequence
of symbols. The MLSE is the optimal method of cancelling the ISI. However, the complexity of
MLSE receiver grows exponentially with the length of the channel impulse response.
Zero-forcing algorithm
The zero-forcing equalizer applies the inverse of the channel to the received signal, to restore the
signal before the channel. It has many useful applications. The name zero forcing corresponds
to bringing down the ISI to zero in a noise-free case. This will be useful when ISI is significant
compared to noise.
For a channel with frequency response F(f ), the zero-forcing equalizer C(f ) is constructed by
C(f ) = 1/F(f ). Thus, the combination of channel and equalizer gives a flat frequency response and
linear phase F(f ) C(f ) = 1.
In reality, zero-forcing equalization does not work in most applications, for the following
reasons:
• Even though the channel impulse response has finite length, the impulse response of the
equalizer needs to be infinitely long.
• The channel may have zeroes in its frequency response that cannot be inverted.
• At some frequencies, the noise may be very small. To compensate, the equalizer amplifies
the noise making it a large value. As a consequence, any noise added after the channel gets
boosted by a large factor and destroys the overall SNR.
on minimization of the MSE between the desired equalizer output and the actual equalizer
output. The LMS algorithm is a linear adaptive filtering algorithm that consists of two basic
processes.
• Filter processes, which involve (1) computing the output of a transversal filter produced
by a set of tap inputs; and (2) generating an estimation error by comparing this output to a
desired response.
• An adaptive process, which involves the automatic adjustment of the tap weights of the
filter in accordance with the estimated error.
Thus, the combination of these two processes working together constitutes a feedback
loop around the LMS algorithm, as illustrated in the block diagram of Figure 14.9(a), and LMS
algorithm filter is shown in Figure 14.9(b).
First, we have a transversal filter, around the LMS algorithm. This component is responsible
for performing the filtering process. Second, we have a mechanism for performing the adaptive
control process on the tap weight of the transversal filter, hence the designation adaptive weight
control mechanism.
∧
u (n) d (n/un)
∧
Transversal filter w (n)
−
Adaptive weight-control e (n) ∑
mechanism
+
d (n)
x (n) Z −1 Z −1 Z −1
WN−1
W0 W1
y (n)
−
e (n) d (n)
+
In practice, the minimization of the MSE is carried out recursively and may be performed by
use of the stochastic gradient algorithm. It is the simplest equalization algorithm and requires
only (2N + 1) operations per iteration. The filter weights are updated by the update equation.
Let variable n denote the sequence of iteration. LMS is computed iteratively by considering
wk ( n ) to denote the tap-weight vector of LMS filter, computed at iteration (time step) n.
The adaptive linear operation of the filter is completely described by the recursive equation
(assuming complex data).
Filter output:
H
y( n ) = w
ˆ (n) x (n − i)
Estimation error:
e( n ) = d( n ) − y( n )
Tap-weight adaptation:
wk ( n + 1) ≡ wk ( n ) + m ek ( n ) x ( n − k ) (14.4)
The LMS equalizer maximizes the signal-to-distortion ratio at its output within the constraints
of the equalizer filter length. If an input signal has a time dispersion characteristic that is greater
than the propagation delay through the equalizer, then the equalizer will be unable to reduce
distortion. The convergence rate of the LMS algorithm is slow due to the fact that there is only
one parameter, the step size that controls the adaptation rate.
In reality, however, exact measurement of the gradient vector is not possible since this
would require prior knowledge of both the correlation matrix R of the tap inputs and the cross-
correlation vector r between the tap inputs and the desired response. Consequently, the gradient
vector must be estimated from the available data. To prevent the adaptation from becoming
unstable, the value of m is chosen from
N
0<m <2 ∑l
i ≡1
i (14.5)
positive number ε is added to the denominator term as the correction factor. Here, the step size
is time varying and is expressed as
b
m (n) = (14.6)
x ( n )) + e
2
where
b is the positive step size
m(n) is the step-size sequence
Therefore, the NLMS algorithm update equation takes the form of
b
wk ( n + 1) = wk ( n ) + ek ( n ) x ( n − k ) (14.7)
x ( n )) + e
2
The NLMS algorithm is utilized, as its convergence speed is, generally, superior to that of
LMS. The NLMS algorithm differs from the LMS algorithm in the incoming reference variable and
is normalized by the magnitude of the incoming reference signal.
There also exists a long-term correlation due to the periodic nature of the speech. This
statistical redundancy can be exploited by introducing prediction schemes, which quantize the
prediction error instead of the speech signal itself. The shortcomings in human capability to
receive sounds, on the other hand, lead to the fact that a lot of information in the speech signal is
perceptually irrelevant. The perceptual irrelevancy means that the human ear cannot differentiate
between changes of magnitude below a certain level and cannot distinguish frequencies below
16 Hz or above 20,000 Hz. This can be exploited by designing optimum quantization schemes,
where only a finite number of levels are necessary.
Speech coding is the process for reducing the bit rate of digital speech representation for
transmission or storage, while maintaining a speech quality that is acceptable for the application.
There are three different speech coding methods, which use different features in different ways:
• Waveform coding
• Source coding
• Hybrid coding
The following sections will explain the speech coding methods in detail. Figure 14.10 shows
the plot between the bit rate (Kbps) on a logarithmic axis and the speech quality classes of “poor
to excellent” corresponding to the five point MOS scale values of 1 to 5, defined by the ITU. It
may be noted that for low complexity and low delay, a bit rate of 32–64 Kbps is required. This
suggests the use of waveform codecs. However, for low bit rate of 4–8 Kbps, hybrid codecs should
be used. These types of codecs tend to be complex with high delay.
Speech
quality
Excellent Waveform
codecs
Hybrid
Good codecs
Fair
Poor
Bad Vocoders
1 2 4 9 16 32 64
Bit rate (Kbps)
of these attributes when considering speech codecs. To have more delay, the codecs with lower
bit rate can be considered as compared to the higher bit rate codecs. They are generally more
complex to implement and often have lower speech quality than the higher bit rate codecs. The
following are the speech codec attributes.
Delay
The delay of a speech codec can have a great impact on its suitability for a particular application.
The components of total system delay include the following:
• Algorithmic delay
• Processing delay
• Transmission delay
Algorithmic delay is mainly due to the delay introduced by frame size, look ahead, and
multiplexing. Most low bit rate speech codecs process a frame of speech data at a time. The
speech parameters are updated and transmitted for every frame. In addition, to analyse the data
properly it is sometimes necessary to analyse data beyond the frame boundary. Hence, before
the speech can be analysed, it is necessary to buffer a frame’s worth of data. The resulting delay
is referred to as algorithmic delay. This delay component cannot be reduced by changing the
implementation, but all other delay components can.
The processing delay is the second major contribution for delay, which comes from the time
taken by the encoder to analyse the speech and the decoder to reconstruct the speech. It depends
on the speed of the hardware used to implement the coder. The sum of the algorithmic and
processing delays is called the one-way codec delay.
The transmission delay, third component of delay, is the time taken for an entire frame of
data to be transmitted from the encoder to the decoder. The total of the three delays (algorithmic,
processing, and transmission delay) is the one-way system delay. In addition, frame interleaving
delay adds an additional frame delay to the total transmission delay. Frame interleaving is
necessary to combat channel fading and is part of the channel coding process.
Complexity
Special purpose hardware is used to implement the speech codecs, such as digital signal processing
(DSP) chips. The following are the main attributes of DSP:
• The computing speed, in millions of instructions per second (MIPS)
• Random-access memory (RAM)
• Read-only memory (ROM)
350 Mobile Cellular Communication
To design a speech codec, the system designer should make estimation about how much
of these resources are to be allocated. Here we have two types of speech codecs based on the
complexity:
• Low complexity codecs, which use less than 15 MIPS
• High complexity codecs, which require 30 MIPS or more
More complexity results in higher costs and greater power usage. For portable applications,
greater power usage means reduced time between battery recharges or use of larger batteries,
which means more expense and weight.
Quality
Quality is the main attribute among all the speech codec attributes. In many applications,
along with the general-purpose noise, that is, thermal noise, shot noise, etc., there is a large
amount of background noise like car noise, street noise, and office noise. How well does the
codec perform under these adverse conditions? What happens when there are channel errors
during transmission? Are the errors detected or undetected? If undetected, the codec must
perform even more robustly than when it is informed that entire frames are in error. How good
does the codec sound when speech is encoded and decoded twice? All these questions must
be carefully evaluated during the testing phase of a speech codec. The speech quality is often
based on the 5-point MOS scale as defined by the International Telecommunication Union-
Technical (ITU-T).
where superscript T denotes matrix transposition and the N-by-1 signal vector x and parameter
vector a are, respectively,
and
a = [a1 , a2 ,..., aN ]
T
(14.10)
The parameters a1, a2,…,aN in effect define the N degrees of freedom available to us in designing
the predictive model. Equation (14.8) readily suggests the tapped delay line (TDL) shown in
Signal Processing in Wireless Systems 351
aN −1 aN
a a
∑ Prediction
∑ ∑
x(t)
Figure 14.11 as the structure for the predictive model. The key question is how do we determine
the filter coefficients? To answer this question, we need a statistical criterion for optimizing the
design of the filter. A criterion widely used in practice is the MSE criterion, defined by
where E denotes the statistical expectation operator and the difference x ( t ) − xˆ ( t ) stands for the
prediction error. It turns out that if x ( t ) is the sample value of a stationary random process, then
the optimum value of the parameter vector a is given by
a = R−1r (14.12)
where the N × N matrix R is the correlation matrix of the tap inputs of the predictive model
and the N × 1 vector r has its elements as the autocorrelation of the input signal x(t) for lags
−1
TS ,2TS ,..., NTS . The symbol R in Equation (14.12) stands for the inverse of the correlation
matrix. If, however, the physical process responsible for the generation of the signal x(t) is non-
stationary (i.e., its statistics vary with time ), then we require the use of an adaptive procedure
whereby the model parameters are allowed to vary with time.
signals are then encoded by using different numbers of bits to fulfil the quality requirements of that
band based on its prominence. The various schemes differ in their accuracies of spectral analysis and
in the bit allocation principle (fixed, adaptive, and semi-adaptive). Two well-known representatives
of this class are sub-band coding (SBC) and adaptive transform coding (ATC).
14.4.5 Vocoders
The source coding is also called a voice codec or vocoders. It is a hardware circuit that converts
the human speech into a digital code and vice versa. Speech coding differs from other forms of
audio coding, as speech is a much simpler signal than most other audio signals, a speech signal
is limited to a bandwidth of 300–3,400 kHz (whereas audio signal is limited to a bandwidth of
0–20 kHz, e.g., audible range). There is a lot of statistical information available about the properties
of speech. The pressure of sound waves radiated from the lips produces the human speech,
although, with some sounds, significant energy emanates also from the nostrils, throat, etc.
In human speech, the air compressed by the lungs excites the vocal cord in two typical
modes. When generating voice sounds, the vocal cord vibrates and generates quasi-periodic voice
sounds. In the case of lower energy unvoiced sounds, the vocal cord does not participate in
voice production and the source acts like a noise generator. The excitation signal is then filtered
through the vocal apparatus, which behaves like a spectral shaping filter. This can be described
adequately by an all pole transfer function that is constituted by the spectral shaping action of
the vocal tract, lip radian characteristics, etc. In the case of vocoders, instead of producing a close
replica of an input signal at the output, an appropriate set of source parameters is generated to
characterize the input signal sufficiently close for a given period of time.
The following steps are used in this process:
The quality of this type of scheme is predetermined by the accuracy of the source model
rather than the accuracy of the quantization of the parameters. The speech quality is limited
by the fidelity of the source model used. The main advantage of vocoders is their low bit rate,
with the penalty of relatively low, synthetic speech quality. Vocoders can be classified into the
frequency-domain and time-domain subclasses. When compared with the time-domain vocoders,
the frequency-domain vocoders are more effective.
bi ci s (t )
Data Channel
source encoder Modulator
Transmit
channel
bˆ i
ˆc i r (t )
Data Channel
sink decoder Demodulator
Interface
Figure 14.12 Block diagram for a digital transmission setup with channel coding
The error correction and detection codes can be broadly classified into three categories:
1. Error correcting capability in terms of the number of errors that it can correct
2. Fast and efficient encoding of the message
3. Fast and efficient decoding of the received message
4. Maximum transfer of information bits per unit time
parity check number, which is normally the product of information bits contained in the block.
Block codes are FEC codes that help to detect and correct a limited number of errors. In a block
encoder, k is the input information bits to the encoder unit and the encoder then uses different
generator polynomials and adds r numbers of extra bits to k bits.
So the total number of output bits from the encoder will be n = r + k. The block code is
referred to as an (n, k) code, and the rate of the code is defined as Rc = k/n. With a k input bit
sequence, 2k distinct code words can be transmitted. For each code word, there is a specific
mapping between the k message bits and the r check bits. The code is systematic because a part
of the sequence in the code word coincides with the k message bits. As a result, it is possible to
make a clear distinction in the code word between the message bits and the parity bits. The code
is binary as these are constructed from bits and linear as each code word can be created by a linear
modulo-2 addition of two or more code words.
To generate an (n, k) block code, the channel encoder accepts information data in successive
k-bit blocks and adds (n − k) redundant bits to each block that are algebraically related to the k
message bits, where k < n, thereby producing an overall encoded block of n bits. The data rate at
which the block encoder produces bits is given by
Ro = ( n / k )Rs (14.13)
Where Ro is the channel data rate at the output of block encoder
Rs is the source information data rate
In GSM cellular communication system, a data block of 180 bits is encoded into 220 bits of code
word on the control channel before sending it to a convolution encoder. Determine the number
of parity check bits added and the code rate of the block encoder used.
Solution
Block codes use algebraic technique properties to encode and decode blocks of data bits or
symbols. The code words are generated in a systematic form such that the original k number of
source information data bits are retained as it is and the (n − k) parity check bits or redundant bits
are either appended or prepended to information data bits. A simple operation of a block code is
depicted in Figure 14.13.
Mostly, operation in a block encoder are based on linear feedback shift registers that are easy
to implement and inexpensive. The parity check bits are generated using a generator polynomial
or matrix. Codes generated by a polynomial are called cyclic codes, and the resultant block codes
are called cyclic redundancy check (CRC) codes. The code word comprising of n bits of encoded
356 Mobile Cellular Communication
Block
encode k-bit data block
k-bit data block
block of data is a transmitter over the channel. The received code word may be error free or
modified due to channel error. Sometimes it may result in another valid code word, which cannot
be detected as error, with a probability given by
• Block codes have limited ability to handle large numbers of distributed errors in a wireless
channel.
• Block codes use hard decisions that tend to destroy information. They do not achieve the
performance as obtained through the use of soft decisions.
• The block encoder accepts a k-bit information data block and generates an n-bit code word.
Thus, code words are generated on a block-by-block data basis. Clearly, a provision must be
made in the block encoder to buffer an entire information data block before generating the
associated code word.
Convertor Output
Combinatory logics parallel
serial
Figure 14.14 General diagram of a convolution encoder with an output of K/N and memory m
the ratio R = K/N is called the code rate. If K information symbols at the encoder input are found
explicitly in the coded block ck, that is,
ck = (dk,1,…,dk,K,ck,K+1,…,ck,N)
Then the code is known as systematic. In the contrary case, it is known as non-systematic. The
general diagram of an encoder with output K/N and memory m is represented in Figure 14.14.
At every moment k, the encoder has m blocks of K information symbols in memory. These m
K binary symbols define the Sk state of the encoder:
Sk = (dk,dk−1,…,dk−(m+1))
If the input of the encoder is permanently fed by blocks of K information symbols, then the
encoder output consists of N infinite sequences of coded symbols, which, for the output i, have
the form:
Let us note that convolutional codes are well adapted to code transmissions with continuous
flow of data. Indeed, the sequences of data to be coded can have any length. To each coded
sequence i = 1, …, N
Generally, convolution coders are more powerful than block codes in terms of providing
FEC, but are not useful for detection or ARQ schemes. At the receiver, FEC is performed using a
maximum likelihood decoding algorithm that determines what data sequence would have been
most likely transmitted, given the received sequence of bits. VLSI implementation of the Viterbi
algorithm is the most common algorithm used for this purpose.
Parity check
Encoder 1 bits z1
Input data
stream Parity check
Interleaver Encoder 2
bits z2
Systematic bits x
Various approaches have been proposed for the construction of powerful codes with large
equivalent block lengths structured in such a way that the decoding can be split into a number
of manageable steps. Building on these previous approaches, the development of turbo codes
and low-density parity-check codes has been by far most successful. Indeed this development
has opened a brand new and exciting way of constructing good codes and decoding them with
feasible complexity.
Turbo coding
In its most basic form, the encoder of a turbo code consists of two constituent systematic encoders
joined together by means of an interleaver, as illustrated in Figure 14.15.
An interleaver is an input output mapping devise that permutes the ordering of a sequence
of symbols from a fixed alphabet in a completely deterministic manner, that is, it takes the
symbols at the input, produces identical symbols at the input, and produces identical symbols at
the output but in a different temporal order. The interleaver can be of many types, of which two
of them are periodic and pseudorandom. Turbo codes use a pseudorandom interleaver, which
operates only on the systematic bits.
There are two reasons for the use of an interleaver in a turbo code:
• To tie together errors that is easily made in one-half of the turbo code to errors that are
exceptionally unlikely to occur in the other half. This is indeed the main reason why the
turbo code performs better than a traditional code.
• To provide robust performance with respect to mismatched decoding, this is a problem that
arises when the channel statistics are not known or have been incorrectly specified.
Turbo decoding
Turbo codes derive their distinctive name from analogy of the decoding algorithm to the turbo
engine principle. Figure 14.16 shows the basic structure of the turbo decoder. It operates on noisy
versions of the systematic bits and two sets of parity check bits in two decoding stages to produce
an estimate of the original message bits.
The first decoder, the interleaver, the second decoder, and the de-interleaver constitute a
single hop feedback system. This arrangement makes it possible to iterate the decoding process in
the receiver many times so as to achieve satisfactory performance. The inputs to the first decoder
are the channel samples corresponding to the systematic information bits, the channel samples
corresponding to the parity bits of the first encoder, and the extrinsic information about the
systematic bits that were determined from the second decoder.
Signal Processing in Wireless Systems 359
De-interleaver
Noisy
Decoder Random Decoder Hard
systematic bits 2 De-interleaver
1 interleave limite
Noisy parity bits x1
Decoded
bits x
Noisy parity bits x2
Each decoding process yields a soft output decision for the next decoder. The key component
is the soft-input, soft-output (SISO) decoder. To achieve the benefits of turbo code, several
iterations are provided. Thus, the turbo codes are processing intensive and are applied to less
delay sensitive applications such as data. They have been implemented in both software and
hardware as a single integrated circuit, and theydo not suffer from the error at low BERs that have
been attributed to other codes. Turbo codes are capable of providing high coding gain, even with
high code rates.
• A turbo code is closer to the random code because it uses a pseudorandom interleaver to
separate its own two convolutional encoders.
• In higher code rates or low SNR value conditions, turbo codes exhibit better performance
than traditional convolutional codes.
• Convolutional codes and turbo codes perform better with soft decisions. However,
convolutional codes can also work with hard decisions.
• Convolutional codes do not have an error floor, whereas turbo codes do have an error floor.
It means that in turbo codes, the BER drops very quickly in the beginning, but eventually
settles down and decreases at a much slower rate.
360 Mobile Cellular Communication
10−1
(9.6 Kbps)
10−3
10−5
Turbo code
(480.8 kbps)
10−6
0 0.5 1.0 1.5 2.0 2.5 3.0 3.5
EB /No(dB)
• Both convolutional and turbo codes require the use of flush bits to initialize them to state
0 at the end of the incoming source information bits sequence. However, due to parallel
encoding structure of turbo codes, it is not straightforward to flush the second encoder.
• Turbo codes are decodable, and hence they are of more practical importance.
Turbo codes are inherently block codes with the block size determined by the size of the
turbo interleaver. These codes are used in 3G cellular technology for high-speed data rate
applications.
14.6 Summary
• In this chapter, different signal processing techniques such as diversity, equalization, speech,
and channel coding are introduced. They have been successfully used in communication
systems to improve the quality of communications.
• Diversity is a commonly used technique in mobile radio systems to combat signal fading to
improve the SNR of the system. The basic principle of diversity, different diversity schemes,
and roles of different diversity signal combining techniques are discussed.
• Equalization is used to overcome ISI due to channel time dispersion.
• Equalization techniques are widely used to improve wireless link performance and received
signal quality.
• Equalization techniques which can combat and/or exploit the frequency selectivity of
the wireless channel are of enormous importance in the design of high data rate wireless
systems.
Signal Processing in Wireless Systems 361
• The efficient utilization of the allocated spectrum is the prime objective in the design of
digital wireless communication systems. The systems rely on the use of speech coding to
remove almost all the natural redundancy inherent in an analogue speech signal, while
ensuring a high quality reproduction of the original speech signal at the receiver.
• Speech coding is the process for reducing the bit rate of digital speech representation for
transmission or storage, while maintaining a speech quality that is acceptable for the
application. Speech coding methods can be classified as waveform coding, source coding,
and hybrid coding.
• Channel coding is a common strategy to make digital transmission more reliable, or,
equivalently, to achieve the same required reliability for a given data rate at a lower power
level at the receiver.
• There are two different approaches for error control coding: ARQ and FEC. ARQ is detection-only
type coding in which transmission errors can only be detected by the receiver but not corrected.
FEC allows not only detection of errors at the receiving end but correction of errors as well.
Review questions
1. What is the basic principle of diversity? Explain different diversity schemes.
2. What is diversity? How is it provided in a communication system?
3. Among the selection, equal gain, and maximal ratio combining, which scheme is the best
and why?
4. What is meant by equalization? Explain adaptive linear equalization.
5. Explain the LMS algorithm.
6. List the main attributes of a speech coding.
7. The channel coding approach which detects and corrects the transmission errors is _______
(a) automatic repeat request (ARQ) (b) forward error correction (FEC)
(c) Both (a) and (b) (d) none
8. Which of the following is the convolutional code _______
(a) Hamming code (b) Viterbi code
(c) Reed–Solomon code (d) BCH code
Answers: 1. (a), 2. (a), 3. (c), 4. (c), 5. (d), 6. (a), 7. (a), 8. (c), 9. (b).
Key equations
1. The coherence bandwidth is given by
1
Bc ≈
Td
2. The adaptive linear operation of the filter is completely described by the recursive equation
wk ( n + 1) ≡ wk ( n ) + m ek ( n ) x ( n − k )
5. The data rate at which the block encoder produces bits is given by
R0 = ( n/k )Rs
Further reading
Atal, B. S. “Predictive Coding of Speech Signals at Low Bit Rates,” IEEE Transactions on
Communications, Vol. 30, No. 4, 1982, pp. 600–614.
Atal, B. S., and Schroeder, M. R. “Stochastic Coding of Speech at Very Low Bit Rate,” Proceedings of
International Conference in Communications, Amsterdam, 1984, pp. 1610–1613.
Signal Processing in Wireless Systems 363
Bahl, L. R., Cocke, J., Jelinek, F., and Raviv, J. “Optimal Decoding of Linear Codes for Minimizing
Symbol Error Rate,” IEEE Transactions on Information Theory, Vol. IT-20, No. 2.
Chen, J., Cox, R., Lin, Y., Jayant, N., and Melchner, M. “Coder for the CCITT 16 kbps Speech
Coder Standard,” IEEE Journal of Selected Areas of Communications, Vol. 6, 1988, pp. 353–363.
Furuskar, A., et al. “System Performance of EDGE: A Proposal for Enhanced Data Rates in Existing
Digital Cellular System,” IEEE VTC 98, pp. 1284–1289.
Hess, W. Pitch Determination of Speech Signals. Berlin: Springer Verlag, 1983.
Jarvinen, K., et al. “GSM Enhanced Full Rate Speech Codec,” IEEE GLOBECOM ’97.
Kleijn, W. B., Ramachandran, R. P., and Kroon, P. “Generalized Analysis by Synthesis Coding
and Its Application to Pitch Prediction,” International Conference on Acoustics. Speech Signal
Processing, San Francisco, 1992, pp. 1337–1340.
Kroon, P., and Deprettere, E. F. “A Class of Analysis by Synthesis Prediction Coders for High
Quality Speech Coding at Rates Between 4.8 and 16 kbps,” IEEE Journal of Selected Areas of
Communications, Vol. 6, 1988, pp. 353–363.