Skip to content
The part of the process nobody writes down
What Nobody ExplainsThe part of the process nobody writes down

Queues & Waiting

What a queue does when the system it depends on fails

An outage does not stop a queue; it converts it into a different kind of queue with different rules, and the rules are usually written down somewhere.

By Lukas Brenner3 min read

Commuters line up at night in Buenos Aires, waiting beside a colorful city bus.
Photograph by Gera Cejas via Pexels
Editorial note. Independent reporting and analysis. Nothing here is sponsored or paid for. How we work.

The line does not disappear when the screens go dark

When the machinery behind a counter stops, the people in front of it stay exactly where they are. This is the awkward property of a physical queue: it has no way to pause, and every minute of the outage adds arrivals who know nothing about it. The queue becomes a container for a problem rather than a mechanism for solving one.

What happens next is more organised than it appears from within the line. Most organisations of any size have a contingency procedure for this, covering what can still be done, what has to be recorded for later and how the queue itself should be managed. The procedure is rarely announced, which is why an outage looks like improvisation from outside.

Fallback is a narrower service, deliberately

The first thing a contingency procedure does is shrink the menu. A handful of transactions are designated as things that can proceed on paper or in a degraded mode, usually the ones that are low value, reversible, or urgent enough that the risk of proceeding is smaller than the risk of waiting.

Everything else stops, and stops for a reason worth understanding. Completing a transaction without the system means completing it without the checks the system performs — without seeing the current balance, the outstanding flag, the note somebody added yesterday. Organisations set the fallback boundary at the point where acting blind stops being safe, and staff at a counter cannot move that boundary because it was not set at the counter.

Manual capture creates work that has to be paid for later

Where paper fallback is permitted, each transaction is written down and keyed in afterwards. That is not a free workaround. Manual entry is slower per item than the original transaction would have been, it is done on top of the following day’s normal workload, and it introduces a second opportunity for error.

It also produces a backlog with an unusual property: the transactions are dated when they occurred but appear in the system when they are entered. For a few days afterwards, records can show sequences that look wrong, balances that update in the wrong order, and confirmations arriving long after the event. None of that means the transaction failed. It means the queue of paper is still being worked through.

Managing the line becomes a separate job from serving it

During a failure the most valuable person in the hall is generally the one not serving anybody. Somebody has to walk the queue, establish who needs a service that is still available, tell the rest plainly that it is not, and get them out of a line that cannot help them.

That is why staff appear to stop working during an outage and start talking to people. Triage is the work. Every person removed from a queue that cannot serve them shortens the wait for everybody who can be served and prevents the far worse outcome, which is somebody standing for forty minutes to reach a counter and being told at the front what could have been said at the back.

Restart is not the moment the queue clears

When service resumes, the queue is longer than it was and the arrival rate has not dropped, so the backlog drains slowly and sometimes not at all before closing. There is also a burst of people returning who left earlier, which arrives on top of everything else.

Recovery therefore depends on capacity that did not exist during the outage: extra positions opened, breaks rescheduled, non-counter staff brought forward. Where an organisation manages an outage well, most of the visible effort happens after the fix rather than during it, and the honest measure of the contingency plan is how the following morning goes rather than how the hour itself looked.

What is worth doing while standing in a stopped queue

Ask which transactions are still possible rather than whether the system is working, because the useful answer is a list rather than a yes or no. Ask whether the request can be lodged now for processing later, which is frequently allowed and rarely offered unprompted.

And if a deadline is in play, ask for the attempt to be recorded with the date and time. Contingency procedures almost always include a way to note that somebody presented themselves during an outage, precisely because the alternative is penalising people for a failure that was not theirs. Obtaining that note while you are standing there takes a minute. Establishing the same fact by correspondence a month later takes considerably longer.

Common questions

Should I wait or come back later?

Ask whether your specific transaction is on the fallback list. If it is, waiting may be worthwhile; if it is not, the queue cannot help you however long you stand in it, and returning after the recovery burst has passed is the faster route.

Why can some people be served and not others?

Because contingency procedures permit a narrow set of transactions, usually the low-risk or urgent ones, and refuse everything requiring checks the system would normally perform. It looks arbitrary from the line and follows a written list from behind the counter.

Will a transaction taken on paper be dated correctly?

Normally yes, since the paper record carries the date and time it was taken and that date is entered along with the transaction. It is still worth keeping any slip you are given, because it is your only evidence until the entry catches up.

Queues & Waitingqueuesoutagescontingencyprocess
Lukas Brenner
Features writer, What Nobody Explains

Lukas has written about behind the counter, paperwork, queues & waiting for most of the last decade and thinks most subjects are more interesting once you know how they work.