Académique Documents
Professionnel Documents
Culture Documents
IJETWORKS
ISDN SYSTEMS
ELSMIER Computer Networks and ISDN Systems 28 (1996) 1723-1738
Abstract
Congestion control mechanisms for ATM networks as selected by the ATM Forum traffic management group are described.
Reasons behind these selections are explained. In particular, selection criterion for selection between rate-based and credit-
based approach and the key points of the debate between these two approaches are presented. The approach that was finally
selected and several other schemes that were considered are described.
Keywords: ATM Networks; Traffic management; Congestion control; flow control; Computer networks; Congestion avoidance
interface (UNI), which specifies how a computer (2) Sustained Cell Rate (SCR) : This is the average
system (which is owned by a user) should commu- rate as measured over a long interval.
nicate with a switch (which is owned by the network (3) Cell Loss Ratio (CLR) : The percentage of cells
service provider). that are lost in the network due to error and congestion
ATM networks are connection-oriented in the sense and are not delivered to the destination.
that before two systems on the network can commu-
Lost Cells
nicate, they should inform all intermediate switches Cell Loss Ratio =
about their service requirements and traffic parame- Transmitted Cells ’
ters. This is similar to the telephone networks where Each ATM cell has a “Cell Loss Priority (CLP)” bit
a circuit is setup from the calling party to the called in the header. During congestion, the network first
party. In ATM networks, such circuits are called virtual drops cells that have CLP bit set. Since the loss of
circuits (VCs) . The connections allow the network to CLP = 0 cell is more harmful to the operation of the
guarantee the quality of service by limiting the num- application, CLR can be specified separately for cells
ber of VCs. Typically, a user declares key service re- with CLP = 1 and for those with CLP = 0.
quirements at the time of connection set up, declares (4) Cell Transfer Delay (CTD) : The delay experi-
the traffic parameters and may agree to control these enced by a cell between network entry and exit points
parameters dynamically as demanded by the network. is called the cell transfer delay. It includes propaga-
This paper is organized as follows. Section 2 defines tion delays, queueing delays at various intermediate
various quality of service attributes. These attributes switches, and service times at queueing points.
help define various classes of service in Section 3. We (5) Cell Delay Variation (CDV) : This is a mea-
then provide a general overview of congestion con- sure of variance of CID. High variation implies larger
trol mechanisms in Section 4 and describe the gen- buffering for delay sensitive traffic such as voice and
eralized cell algorithm in Section 4.1. Section 5 de- video. There are multiple ways to measure CDV. One
scribes the criteria that were set up for selecting the measure called “peak-to-peak” CDV consists of com-
final approach. A number of congestion schemes are puting the difference between the (1 - a)-percentile
described briefly in Section 6 with the credit-based and the minimum of the cell transfer delay for some
and rate-based approaches described in more details small value of (Y.
in Sections 7 and 8. The debate that lead to the even- (6) Cell Delay Variation Tolerance (CDVT) and
tual selection of the rate-based approach is presented Burst ToZerance (B7’) : For sources transmitting at any
in Section 9. given rate, a slight variation in the inter-cell time is
The description presented here is not intended to be allowed. For example, a source with a PCR of 10,000
a precise description of the standard. In order to make cells per second should nominally transmits cells every
the concept easy to understand, we have at times sim- 100~s. A leaky bucket type algorithm called “Gener-
plified the description. Those interested in precise de- alized Cell Rate Algorithm (GCRA)” is used to de-
tails should consult the standards documents, which termine if the variation in the inter-cell times is ac-
are still being developed as of this writing in Jan- ceptable. This algorithm has two parameters. The first
uary 1995. parameter is the nominal inter-cell time (inverse of the
rate) and the second parameter is the allowed variation
in the inter-cell time. Thus, a GCRA( 100,~s lops),
will allow cells to arrive no more than 10 ,us earlier
2. Quality of Service (QoS) and traffic attributes than their nominal scheduled time. The second param-
eter of the GCRA used to enforce PCR is called Cell
While setting up a connection on ATM networks, Delay Variation Tolerance (CDVT) and of that used
users can specify the following parameters related to to enforce SCR is called Burst Tolerance (BT)
the input traffic characteristics and the desired quality (7) Maximum Burst Size (MBS): The maximum
of service. number of back-to-back cells that can be sent at the
( 1) Peak Cell Rate (PCR) : The maximum instan- peak cell rate but without violating the sustained cell
taneous rate at which the user will transmit. rate is called maximum burst size (MBS) . It is related
R. Jain/Computer Networks and ISDN Systems 28 (I 996) 1723-l 738 1725
Table I
ATM layer service categories
Attribute CBR rt-VBR rut-VBR UBR ABR
PCR and CDVT Specified Specified Specified SpecifiedC Spccifiedd
SCR, MBS, CDVTe,’ n/a Specified Specified n/a n/a
MCR’ n/a n/a n/a n/a Specified
Peak-to-peak CDV Specified Specified Unspecified Unspecified Unspecified
Mean CTD Unspecified Unspecified Specified Unspecified Unspecified
Maximum CTD Specified Specified Unspecified Unspecified Unspecified
CLR= Specified Specified Specified Unspecified Specifiedb
Feedback Unspecified Unspecified Unspecified Unspecified Specified
a For all service categories, the cell loss ratio may be unspecified for CLP = 1.
b Minimized for sources that adjust cell flow in response to control information.
c May not be subject to connection admission control and user parameter control procedures.
d Represents the maximum rate at which the source can send as controlled by the control information.
e These parameters are either explicitly or implicitly specified.
r Different values of CDVT may specified for SCR and PCR.
schemes.Typically, a combination of such schemes For very short spikesin traffic load, providing suf-
is used. The selection dependsupon the severity and ficient buffers in the switches is the best solution.
duration of congestion. Notice that solutions that are good for short term
Fig. 1 showshow the duration of congestionaffects congestion are not good for long-term overload and
the choice of the method. The best method for net- vice-versa. A combination of various techniques
works that are almost always congested is to install (rather than just one technique) is used since over-
higher speedlinks and redesignthe topology to match loads of various durations are experienced on all
the demand pattern. networks.
For sporadic congestion, one method is to route ac- UNI 3.0 allows CAC, traffic shaping, and binary
cording to load level of links and to reject new con- feedback (EFCI). However, the algorithms for CAC
nections if all paths are highly loaded. This is called are not specified. The traffic shaping and feedback
“connection admission control (CAC)“. The “busy” mechanismsare described next.
tone on telephone networks is an example of CAC.
CAC is effective only for medium duration congestion
since once the connection is admitted the congestion
may persist for the duration of the connection. 4.1. Generalized Cell Rate Algorithm (GCRA)
For congestions lasting less than the duration of
connection, an end-to-end control schemecan be used. As discussedearlier, GCRA is the so-called “leaky
For example, during connection setup, the sustained bucket” algorithm, which is used to enforce regular-
and peak rate may be negotiated. Later a leaky bucket ity in the cell arrival times. Basically, all arriving cells
algorithm may be usedby the sourceor the network to are put into a bucket, which is drained at the speci-
ensurethat the input meetsthe negotiatedparameters. fied rate. If too many cells arrive at once, the bucket
Such “traffic shaping algorithms” are open loop in the may overflow. The overflowing cells are called non-
sensethat the parameterscannot be changed dynam- conforming and may or may not be admitted in to the
ically if congestion is detected after negotiation. In a network. If admitted, the cell loss priority (CLP) bit
closedloop scheme,on the other hand, sourcesare in- of the non-conforming cells may be set so that they
formed dynamically about the congestion state of the will be first to be dropped in caseof overload. Cells
network and are asked to increase or decreasetheir of service categoriesspecifying peak cell rate should
input rate. The feedback may be usedhop-by-hop (at conform to GRCA( 1/PCR, CDVT), while those also
datalink layer) or end-to-end (transport layer). Hop- specifying sustainedcell rate should additionally con-
by-hop feedback is more effective for shorter term form to GCRA( l/SCR, BT) . See Section 2 for deti-
overloads than the end-to-end feedback. nitions of CDVT and BT.
R. -lain/Computer Networks and ISDN Systems 28 (1996) 1723-I 738 1727
UNI V3.0 specified two different facilities for feed- of general interest and apply to non-ATM networks
back control: as well, we describe some of them briefly here.
( 1) Generalized Flow Control (GFC) .
(2) Explicit Forward Congestion Indication (EFCI) . 5.1, Scalability
As shown in Fig. 2, the first four bits of the cell
header at the user-network interface (UNI) were re- Networks are generally classified based on extent
served for GFC. At network-network interface, GFC (coverage), number of nodes, speed, or number of
is not used and the four bits are part of an extended users. Since ATM networks are intended to cover a
virtual path (VP) field. It was expected that the GFC wide range along all these dimensions, it is necessary
bits will be used by the network to flow control the that the scheme be not limited to a particular range
source. The GFC algorithm was to be designed later. of speed, distance, number of switches, or number of
This approach has been abandoned. VCs. In particular, this ensures that the same scheme
Switches can use the payload type field in the cell can be used for local area networks (LANs) as well
header to convey congestion information in a binary as wide area networks ( WANs) .
(congestion or no congestion) manner. When the first
bit of this field is 0, the second bit is treated as “ex- 5.2. Optima&y
plicit forward congestion indication (EFCI) bit”. For
all cells leaving their sources, the EFCI bit is clear. In a shared environment the throughput for a source
Congested switches on the path, if any, can set the depends upon the demands by other sources. The most
EFCI bit. The destinations can monitor EFCI bits and commonly used criterion for what is the correct share
ask sources to increase or decrease their rate. The ex- of bandwidth for a source in a network environment, is
act algorithm for use of EFCI was left for future defi- the so-called “max-min allocation [ 121”. It provides
nition. The use of EFCI is discussed later in Section 8. the maximum possible bandwidth to the source receiv-
ing the least among all contending sources. Mathemat-
ically, it is defined as follows. Given a configuration
5. Selection criteria with n contending sources, suppose the ith source gets
a bandwidth xi. The allocation vector {xl, x2, . . . , x”}
ATM network design started initially in CCIIT is feasible if all link load levels are less than or equal to
(now known as ITU). However, the progress was 100%. The total number of feasible vectors is infinite.
rather slow and also a bit “voice-centric” in the sense For each allocation vector, the source that is getting
that many of the decisions were not suitable for data the least allocation is in some sense, the “unhappiest
traftic. So in October 1991, four companies - Adap- source”. Given the set of all feasible vectors, find the
tive (NET), CISCO, Northern Telecom, and Sprint, vector that gives the maximum allocation to this un-
formed ATM Forum to expedite the process. Since happiest source. Actually, the number of such vectors
then ATM Forum membership has grown to over 200 is also infinite although we have narrowed down the
principal members. The traflic management working search region considerably. Now we take this “unhap-
group was started in the Forum in May 1993. A num- piest source” out and reduce the problem to that of
ber of congestion schemes were presented. To sort remaining n - 1 sources operating on a network with
out these proposals, the group decided to first agree reduced link capacities. Again, we find the unhappi-
on a set of selection criteria. Since these criteria are est source among these n - 1 sources, give that source
I728 R. JaidComputer Networks rind IS&V Systems 28 (1996) 1723-l 738
the maximum allocation and reduce the problem by Given any optimality criterion, one can determine
one source. We keep repeating this process until all the optimal allocation. If a scheme gives an alloca-
sources have been given the maximum that they could tion that is different from the optimal, its unfairness is
get. quantified numerically as follows. Suppose a scheme
The following example illustrates the above concept allocates { Zt , &, . . . , a,,} instead of the optimal allo-
of max-min fairness. Fig. 3 shows a network with four cation {at, 2.2, . . . , a,}. Then, we calculate the nor-
switches connected via three 150Mbps links. Four malized allocations xi = ii/ii for each source and
VCs are setup such that the first link Ll is shared by compute the fairness index as follows [ 13,141:
sources S 1, S2, and S3. The second link is shared by S3
and S4. The third link is used only by S4. Let us divide
the link bandwidths fairly among contending sources.
On link Ll , we can give 50Mbps to each of the three Since allocations Xis usually vary with time, the fair-
contending sources Sl, S2, and S3. On link L2, we ness can be plotted as a function of time. Alternatively,
would give 75 Mbps to each of the sources S3 and S4. throughputs over a given interval can be used to com-
On link L3, we would give all 155 Mbps to source pute overall fairness.
S4. However, source S3 cannot use its 75Mbps share
at link L2 since it is allowed to use only 50Mbps at
5.4. Robustness
link Ll. Therefore, we give 50 Mbps to source S3 and
construct a new configuration shown in Fig. 4, where
The scheme should be insensitive to minor devia-
Source S3 has been removed and the link capacities
tions. For example, slight mistuning of parameters or
have been reduced accordingly. Now we give l/2 of
loss of control messages should not bring the network
link Ll’s remaining capacity to each of the two con-
down. It should be possible to isolate misbehaving
tending sources: S1 and S2; each gets 50Mbps. Source
users and protect other users from them.
S4 gets the entire remaining bandwidth (1OOMbps)
of link L2. Thus, the fair allocation vector for this con-
figuration is (50, 50, 50, 100). This is the max-min 5.5. Implementability
allocation.
Notice that max-min allocation is both fair and ef- The scheme should not dictate a particular switch
ficient. It is fair in the sense that all sources get an architecture. As discussed later in Section 9, this
equal share on every link provided that they can use turned out to be an important point in final selection
it. It is efficient in the sense that each link is utilized since many schemes were found to not work with
to the maximum load possible. FIFO scheduling.
It must be pointed out that the max-min fairness is
just one of several possible optimality criteria. It does 5.6. Simulation configurations
not account for the guaranteed minimum (MCR).
Other criterion such as weighted fairness have been A number of network configuration were also
proposed to determine optimal allocation of resources agreed upon to compare various proposals. Most
over and above MCR. of these were straightforward serial connection of
R. Jain/Computer Networks and ISDN Systems 28 (1996) 1723-l 738 1719
6. Congestion schemes
Although the proposal was presented at the ATM and rate-based proposals being discussed at the time.
Forum, it was not followed up and the precise details It consists of using window flow control on every link
of how the delay will be used were not presented. and to use binary (EFCI-based) end-to-end rate con-
Also, this method does not really require any stan- trol. The window control is per-link (and not per-VC
dardization, since any source-destination pair can do as in credit-based scheme). It is, therefore, scalable
this without involving the network. in terms of number of VCs and guarantees zero cell
loss. Unfortunately, neither the credit-based nor the
6.3. Backward Explicit Congestion NotiJcation rate-based camp found it acceptable since it contained
(BECN) elements from the opposite camp.
This method presented by N.E.T. [ 28-301 consists 4.6. Fair queueing with rate and bufer feedback
of switches monitoring their queue length and sending
an RM cell back to source if congested. The sources This proposal from Xerox and CISCO [ 271 consists
reduce their rates by half on the receipt of the RM of sources periodically sending RM cells to determine
cell. If no BECN cells are received within a recovery the bandwidth and buffer usage at their bottlenecks.
period, the rate for that VC is doubled once each period The switches compute fair share of VCs. The mini-
until it reaches the peak rate. To achieve fairness, the mum of the share at this switch and that from previ-
source recovery period was made proportional to the ous switches is placed in the RM cells. The switches
VC’s rate so that lower the transmission rate the shorter also monitor each VC’s queue length. The maximum
the source recovery period. of queue length at this switch and those from the pre-
This scheme was dropped because it was found to be vious switches is placed in the same RM cell. Each
unfair. The sources receiving BECNs were not always switch implements fair queueing, which consists of
the ones causing the congestion [ 321. maintaining a separate queue for each VC and com-
puting the time at which the cell would finish trans-
6.4. Early packet discard mission if the queues were to be served round-robin
one-bit at a time. The cells are scheduled to transmit
This method presented by Sun Microsystems [ 351 in this computed time order.
is based on the observation that a packet consists of The fair share of a VC is determined as the inverse
several cells. It is better to drop all cells of one packet of the interval between the cell arrival and its trans-
then to randomly drop cells belonging to different mission. The interval reflects the number of other VCs
packets. In AALS, when the first bit of the payload that are active. Since the number and hence the inter-
type bit in the cell header is 0, the third bit indicates val is random, it was recommended that the average
“end of message (EOM)“. When a switch’s queues of several observed interval be used.
start getting full, it looks for the EOM marker and it This scheme requires per-VC (fair) queueing in the
drops all future cells of the VC until the “end of mes- switches, which was considered rather complex.
sage” marker is seen again.
It was pointed out [ 331 that the method may not be
fair in the sense that the cell to arrive at a full buffer 7. Credit-based approach
may not belong to the VC causing the congestion.
Note that this method does not require any inter- This was one of the two leading approaches and
switch or source-switch communication and, there- also the first one to be proposed, analyzed, and imple-
fore, it can be used without any standardization. Many mented. Originally proposed by Professor H.T. Kung,
switch vendors are implementing it. it was supported by Digital, BNR, FORE, Ascom-
Timeplex, SMC, Brooktree, and Mitsubishi [ 26,411.
6.5. Link window with end-to-end binary Rate The approach consists of per-link, per-VC, window
flow control. Each link consists of a sender node
This method presented by Tzeng and Siu [ 451, con- (which can be a source end system or a switch) and a
sisted of combining good features of the credit-based receiver node (which can be a switch or a destination
R. JainlComputer Networks and ISDN Systems 28 (1996) 1723-I 738 1731
rent rate of the VC in addition to its congestion level Charny during her master thesis work at the Mas-
in deciding whether to set the EFCI in the cell. The sachusettsInstitute of Technology (MIT) [ 51. The
switch computesa “fair share” and if congestedit sets MIT schemeconsistsof each source sending an RM
EFCI bits in cells belonging to only thoseVCs whose cell every nth data cell. The RM cell contains the
rates are above this fair share.The VCs with rates be- VC’s current cell rate (CCR) and a “desired rate”.
low fair share are not affected. The switches monitor all VC’s rates and compute a
“fair share”. Any VCs whosedesired rate is lessthan
8.1. The MIT scheme the fair shareis granted the desired rate. If a VC’s de-
sired rate is more than the fair share,the desired rate
In July 1994, we [6] argued that the binary feed- field is reduced to the fair shareand a “reduced bit” is
back was too slow for rate-basedcontrol in high-speed set in the RM cell. The destination returns the RM cell
networks and that an explicit rate indication would back to the source, which then adjusts its rate to that
not only be faster but would offer more flexibility to indicated in the RM cell. If the reduced bit is clear, the
switch designers. sourcecould demanda higher desiredrate in the next
The single-bit binary feedback can only tell the RM cell. If the bit is set, the source use the current
source whether it should go up or down. It was de- rate as the desiredrate in the next RM cell.
signed in 1986 for connectionlessnetworks in which The fair shareis computed using a iterative proce-
the intermediate nodes had no knowledge of flows or dure as follows. Initially, the fair share is set at the
their demands.The ATM networks are connection ori- link bandwidth divided by the number of active VC’s.
ented. The switches know exactly who is using the All VCs, whose rates are less than the fair share are
resourcesand the flow paths are rather static. This in- called “underloading VCs”. If the number of under-
creasedinformation is not usedby the binary feedback loading VCs increasesat any iteration, the fair share
scheme. is recomputed as follows:
Secondly and more importantly, the binary feed-
Fair Share
back schemeswere designed for window-basedcon-
Link Bandwidth - c Bandwidth of Underloading VCs
trols and are too slow for rate-basedcontrols. With =
Number of VCs - Number of Underloading VCs
window-basedcontrol a slight difference between the
current window and the optimal window will show up The iteration is then repeateduntil the number of un-
as a slight increasein queue length. With rate-based derloading VCs and the fair share does not change.
control, on the other hand, a slight difference in cur- Charny [5] has shown that two iterations are suffi-
rent rate and the optimal rate will show up as contin- cient for this procedure to converge.
uously increasing queue length [ 15,161. The reaction Chamy also showed that the MIT schemeachieve
times have to be fast. We can no longer afford to take max-min optimality in 4k round trips, where k is the
several round trips that the binary feedback requiresto number of bottlenecks.
settle to the optimal operation. The explicit rate feed- This proposal was well received except that the
back can get the sourceto the optimal operating point computation of fair sharerequires order rr operations,
within a few round trips. where n is the number of VCs. Search for an O( 1)
The explicit rate schemeshave several additional schemeled to the EPRCA algorithm discussednext.
advantages.First, policing is straightforward. The en-
try switches can monitor the returning RM cells and 8.2. Enhanced PRCA (EPRCA)
use the rate directly in their policing algorithm. Sec-
ond with fast convergence time, the system come to The merger of PRCA with explicit rate schemelead
the optimal operating point quickly. Initial rate has to the “Enhanced PRCA (EPRCA)” schemeat the
lessimpact. Third, the schemesare robust againster- end of July 1994 ATM Forum meeting [ 36,381. In
rors in or loss of RM cells. The next correct RM cell EPRCA, the sourcessenddata cells with EFCI set to
will bring the system to the correct operating point. 0. After every n data cells, they send an RM cell.
We substantiatedour argumentswith simulation re- The RM cells contain desired explicit rate (ER),
sults for an explicit rate scheme designed by Anna current cell rate (CCR) , and a congestion indication
R. Jain/Computer Networks and ISDN Systems 28 (1996) 1723-l 738 1733
(CI) bit. The sources initialize the ER field to their 8.3. OSU time-based congestion avoidance
peak cell rate (PCR) and set the CI bit to zero.
The switches compute a fair share and reduce the Jain, Kalyanaraman, and Viswanathan at the Ohio
ER field in the returning RM cells to the fair share State University (OSU) have developed a series of
if necessary. Using exponential weighted averaging a explicit rate congestion avoidance schemes.The first
mean allowed cell rate (MACR) is computed and the scheme [ 19,201 called the OSU schemeconsistsof
fair share is set at a fraction of this average: switchesmeasuringtheir input rate over a fixed “aver-
aging interval” and comparing it with their target rate
MACR = ( 1 - a)MACR + &CR, to compute the current Load factor Z:
Fair Share = SWDPF x MACR. Input Rate
Load Factor z =
Here, a is the exponential averaging factor and Target Rate ’
SWDPF is a multiplier (called switch down pressure The target rate is set at slightiy below, say, 85-95s
factor) set close to but below I. The suggested values of the link bandwidth. Unless the load factor is close
of LYand SWDPF are l/ 16 and 7/8, respectively. to 1, all VCs are askedto change (divide) their load
The destinationsmonitor the EFCI bits in data cells. by this factor Z. For example, if the load factor is 0.5.
If the last seen data cell had EFCI bit set, they mark all VCs are asked to divide their rate by a factor of
the CI bit in the RM cell. 0.5, that is, double their rates. On the other hand, if
In addition to setting the explicit rate, the switches the load factor is 2, all VCs would be asked to halve
can also set the CI bit in the returning RM cells if their their rates.
queue length is more than a certain threshold. Note that no selective feedback is taken when the
The sourcesdecreasetheir rates continuously after switch is either highly overloaded or highly under-
every cell. loaded. However, if the load factor is closeto one, be-
ACR = ACR x RDF. tween I - A and 1 + A for a small A, the switch gives
different feedback to underloading sourcesand over-
Here, RDF is the reduction factor. When a source re- loading sources.A fair shareis computed as follows:
ceives the returned RM cell, it increasesits rate by an
Target Rate
amount AIR if permitted. Fair Share=
Number of Active Sources’
IfCI=O
All sources,whose rate is more than the fair shareare
Then New ACR = Min( ACR + AIR, ER, PCR) .
askedto divide their rates by z/( 1 + A) while those
If CI bit is set, the ACR is not changed. below the fair shareare askedto divide their rates by
Notice that EPRCA allows both binary-feedback z/( 1 - A). This algorithm called “Target Utilization
switches and the explicit feedback switches on the Band (TUB) algorithm” was proven to lead to fair-
path. The main problem with EPRCA as described ness[19].
here is the switch congestion detection algorithm. It is The OSU schemehasthree distinguishing features.
basedon queue length threshold. If the queue length First, it is a congestion avoidance scheme. It gives
exceeds a certain threshold, the switch is said to be high throughput and low delay. By keeping the target
congested. If it exceed another higher threshold, it rate slightly below the capacity, the algorithm ensures
said to be very highly congested.This method of con- that the queues are very small, typically close to 1,
gestion detection was shown to result in unfairness. resulting in low delay. Second,the switcheshave very
Sources that start up late were found to get lower few parameterscompared to EPRCA and are easy to
throughput than those which start early. set. Third, the time to reach the steady state is very
The problem was fixed by changing to queue small. The sourcereach their final operating point 10
growth rate as the load indicator. The change in the to 20 times faster than that with EPRCA.
queue length is noted down after processing, say, K In the original OSU scheme, the sources were re-
cells. The overload is indicated if the queue length quired to send RM cells periodically at fixed time in-
increases[ 44,431. terval. This meantthat the RM cell overheadper source
1734 R. JainIComputer Networks and ISDN Systems 28 (1996) 1723-I 738
for the credit-based schemes were found to be less failed. The rate-based approach won by a vote of 7
than those in the rate-based scheme with binary feed- to 104.
back. However, this advantage disappeared when ex-
plicit rate schemes were added. In the static credit-
based approach, per-VC buffer requirement is propor- IO. Summary
tional to link delay, while in the rate-based approach,
total buffer requirement is proportional to the end-to- Congestion control is important in high speed net-
end delay. The adaptive credit scheme allows lower works. Due to larger bandwidth-distance product, the
buffering requirements than the static scheme but at amount of data lost due to simultaneous arrivals of
the cost of increased ramp-up time. bursts from multiple sources can be larger. For the suc-
Note that the queueing delays have to be added in cess of ATM, it is important that it provides a good
both cases since it delays the feedback and adds to the traffic management for both bursty and non-bursty
reaction time. sources.
(6) Delay Estimate: Setting the congestion con- Based on the type of the traffic and the quality of
trol parameters in the credit-based approach requires service desired, ATM applications can use one of the
knowledge of link round trip delay. At least, the link five service categories: CBR, rt-VBR, nrt-VBR, UBR,
length and speed must be known. This knowledge is and ABR. Of these, ABR is expected to be the most
not required for rate-based approaches (although it commonly used service category. It allows ATM net-
may be helpful). works to control the rates at which delay-insensitive
(7) Switch Design Flexibility: The explicit rate data sources may transmit. Thus, the link bandwidth
schemes provide considerable flexibility to switches not used by CBR and VBR applications can be fairly
in deciding how to allocate their resources. Different divided among ABR sources.
switches can use different mechanisms and still in- After one year of intense debate on ABR traffic
teroperate in the same network. For example, some management, two leading approaches emerged: credit-
switches can opt for minimizing their queue length, based and rate-based. The credit-based approach uses
while the others can optimize their throughput, while per-hop per-VC window flow control. The rate-based
still others can optimize their profits. On the other approach uses end-to-end rate-based control. The rate-
hand, the credit-based approach dictated that each based approach was eventually selected primarily be-
switch use per-VC queueing with round-robin service. cause it does not require all switches to keep per-
(8) Switch vs End-System Complexity: The credit- VC queueing and allows considerable flexibility in re-
based approach introduces complexity in the switches source allocation.
but may have made the end-system’s job a bit simpler. Although, the ATM Forum traffic management ver-
The proponents of credit-based approach argued that sion 4.0 specifications allows older EFCI switches
their host network interface card (NIC) is much sim- with binary feedback for backward compatibility, the
pler since they do not have to schedule each and every newer explicit rate switches will provide better perfor-
cell. As long as credits are available, the cells can be mance and faster control. The ATM Forum has speci-
sent at the peak rate. The proponents of the rate-based fied a number of rules that end-systems and switches
approach countered that all NIC cards have to have follow to adjust the ABR traffic rate. The specifica-
schedulers for their CBR and VBR traffic and using tion is expected to be finalized by February 1996 with
the same mechanism for ABR does not introduce too products appearing in mid to late 1996.
much complexity.
There were several other points that were raised.
But all of them are minor compared to the per-VC Acknowledgement
queueing required by the credit-based approach. Ma-
jority of the ATM Forum participants were unwilling Thanks to Dr. Shirish Sathaye of FORE Systems,
to accept any per-VC action as a requirement in the Dr. David Hughes of StrataCorn, and Rohit Goyal and
switch. This is why, even an integrated proposal al- Ram Viswanathan of OSU for useful feedback on an
lowing vendors to choose either of the two approaches earlier version of this paper.
R. Jain/Computer Networks and ISDN Systems 28 (1996) 1723-l 738 1731