John Little proved it in 1961 and it holds for any stable system, regardless of how the work arrives or in what order it is handled. That generality is what makes it useful.

L = λW. The average number of items in the system equals the arrival rate multiplied by the average time each item spends in it. Rearranged, the form that matters operationally:

Cycle time = work in progress ÷ throughput.

What it says

You have twenty active client projects. You finish two a week. Then the average project takes ten weeks from start to finish, and no amount of urgency about any individual one changes that, because the arithmetic does not care about effort.

To shorten cycle time there are exactly two levers. Increase throughput, which means more capacity and costs money. Or reduce work in progress, which is free and is almost never done.

Cut to ten active projects at the same two a week and the average project now takes five weeks. Nobody worked faster. The same work completed in the same total time, and each individual item spent half as long in the system.

Why this is counterintuitive

Starting work feels like progress and an idle person feels like waste, so the instinct is to begin more things. That raises work in progress, which lengthens everything, while throughput stays at whatever capacity allows.

It also has a compounding cost, because items in progress are not free to hold: each one carries context-switching, status updates, and a client waiting and asking. High work in progress generates coordination load that reduces the throughput that was already the constraint.

Using it

Measure your three numbers. Count active items, count completions per week, divide. Most people are surprised by the result and it takes ten minutes.

Set a limit on work in progress and hold it. Nothing starts until something finishes. This is the whole mechanism of a kanban board and it is the cheapest operational improvement available to most businesses.

Resist starting work to look busy. Capacity is measured by what finishes, not by what is open.

It also identifies where to look when cycle time is bad, which is the same place the bottleneck points: throughput is set by the constraint, so adding work anywhere else only lengthens queues.