Topics
Suppose a sequence of events is separated by independent intervals. A report lists the observed intervals and gives their arithmetic mean. It is tempting to interpret that mean as the duration a randomly arriving observer is likely to experience. Renewal theory shows that this interpretation is generally wrong, because choosing an interval uniformly from the event sequence and choosing a random time on the clock are different sampling experiments.
Long intervals occupy more calendar time than short intervals. A ten-hour gap creates ten times as many opportunities for a random observation time to land inside it as a one-hour gap. The interval containing a random time is therefore not distributed like an ordinary inter-arrival time. It is length biased.
This simple fact produces the inspection paradox, the waiting-time paradox and several closely related sampling effects in reliability, queueing, survival analysis and prevalence studies. The paradox is not a contradiction. It is a change of measure created by the observation scheme.
The cleanest way to see the mechanism is through a renewal process.
A renewal process separates event time from observation time
Let
be independent positive inter-arrival times with common distribution $F$ and finite mean
Define renewal epochs
with
The counting process
records how many renewals have occurred by time $t$.
If an analyst chooses an interval by choosing an event index $i$ uniformly from a long list of renewals, the sampled duration behaves like
Its mean is simply
Now consider a different experiment. Observe the system at a time selected uniformly from a very long calendar window, independently of the renewal process, and record the interval containing that observation time.
An interval of length $x$ contributes $x$ units of calendar time to the observation window. All else equal, it is therefore $x$ times as likely to be selected by time sampling as by event sampling.
The probability law changes.
This distinction is the entire inspection paradox.
Random-time sampling produces a size-biased interval distribution
Let $L$ denote the length of the renewal interval containing a random stationary observation time.
For a continuous distribution with density $f$, the density of $L$ is
More generally, without requiring a density, the size-biased law can be written as
The normalizing constant is the ordinary mean interval because
The expected observed interval length is therefore
Writing
we obtain
Hence
with strict inequality whenever the interval distribution is non-degenerate.
The inflation is exactly
The inspection paradox is therefore not caused by the mean being poorly estimated. Even if the event-level interval distribution is known perfectly, time sampling targets a different distribution.
Variance determines how different the two sampling schemes become.
A process with mean interval 1.9 hours is usually observed inside a 10-hour gap
Consider an intentionally simple renewal process:
and
Ninety percent of intervals last one hour. Ten percent last ten hours.
The ordinary mean interval is
hours.
If we choose one interval uniformly from the event sequence, the probability of selecting a ten-hour interval is only
A random observation time sees something very different.
Under size-biased sampling,
while
Although only one interval in ten is long, a random observer is more likely than not to land inside a ten-hour interval.
The reason is mechanical. In every 19 hours of expected renewal time, the short intervals contribute about nine hours and the long intervals contribute about ten hours. Calendar time is divided nearly evenly between the two interval types even though event counts are not.
The second moment is
The mean length of the interval seen at a random time is therefore
hours.
The ordinary event-average interval is 1.9 hours.
The randomly observed interval averages 5.74 hours.
Nothing about the process changed. Only the sampling scheme changed.
Age and residual life split the observed interval
At observation time $t$, define the age of the current interval,
and the residual life,
The enclosing interval length is
In an equilibrium renewal process, the observation time is uniformly distributed within the selected interval conditional on its length.
Given
the expected age and residual life are both
Averaging over the size-biased interval distribution gives
Hence, provided the second moment is finite,
For the two-point example,
hours.
A new interval drawn at a renewal epoch has expected total length 1.9 hours.
A random observer entering at an arbitrary time expects to wait 2.87 hours for the next renewal.
The forward wait is longer than the ordinary mean interval in this example because the interval distribution is highly variable.
This is the classical waiting-time paradox.
The same identity can be expressed through the coefficient of variation.
Let
Because
the equilibrium residual mean is
This formula adds an important qualification to the usual informal story.
The expected residual wait exceeds the ordinary mean interval only when
It equals the mean when
and it is smaller when
Length bias always makes the enclosing interval longer on average when there is any variability.
It does not always make the forward wait exceed the ordinary mean interval.
Those are different statements.
The equilibrium residual-life distribution is an integrated tail
The residual-life distribution can be written without first conditioning on the enclosing interval.
Let
be the survival function of the inter-arrival time.
In equilibrium,
The residual-life distribution is therefore proportional to the integrated tail of the original interval distribution.
Differentiating when $F$ has a density gives the equilibrium residual density
This formula explains why long tails are amplified.
If
decays slowly, substantial residual probability remains at large $x$ because the observer is preferentially located inside long intervals.
The mean follows by integrating the survival function:
Substituting the equilibrium tail gives
Interchanging the order of integration yields
Using the standard identity
we recover
The second moment appears because random-time observation weights duration by duration.
Long intervals influence the result twice: once because they are longer and therefore more likely to be sampled, and again because, once selected, they leave potentially long residual waits.
Deterministic intervals, Poisson intervals and variable intervals behave differently
The coefficient-of-variation form makes three important cases transparent.
If every interval is exactly
then
A random observer is uniformly located somewhere inside a deterministic interval, so
There is no length bias in the enclosing interval because every interval has the same duration.
The forward wait is half the interval on average.
Now suppose inter-arrival times are exponential with rate
Then
and
The residual-life formula gives
More strongly, the residual life itself is exponential with the same rate:
This is the memoryless property.
Arriving at an arbitrary time is statistically equivalent, for the forward wait, to starting immediately after an event.
The enclosing interval is still length biased. Because
we have
The average interval containing a random time is twice the ordinary mean interval.
The Poisson process removes the waiting-time paradox only in the forward direction.
It does not remove length bias in the total enclosing interval.
Finally, when
as in bursty or heavy-tailed inter-arrival processes, the expected residual wait exceeds the ordinary mean interval.
The more variable the intervals, the more misleading the event-level mean becomes for a random observer.
The elementary renewal theorem governs the long-run event rate
The inspection paradox concerns what a random time sees.
Renewal theory also provides the long-run event frequency.
Under standard assumptions with
the elementary renewal theorem states that
as
A stronger almost-sure renewal law gives
under familiar iid conditions.
The reciprocal mean interval is therefore the long-run event rate.
For the two-point example,
renewals per hour.
This rate is perfectly consistent with the random observer seeing long intervals.
Event frequency and time occupancy answer different questions.
A process can generate many short intervals numerically while a large share of calendar time is spent inside a relatively small number of long intervals.
This distinction is central in reliability.
Suppose failures are followed by operational periods of highly variable length. The average interval between failure epochs describes event frequency.
A randomly chosen inspection time describes occupancy of operational intervals in proportion to how long they last.
Confusing the two leads directly to duration bias.
Prevalence sampling is renewal length bias in another language
The same mathematics appears in survival studies.
Imagine episodes with durations
An incidence sample recruits episodes when they begin.
A prevalence sample observes the population at a random calendar time and recruits episodes currently in progress.
Long episodes are more likely to be active at the sampling time.
The observed duration distribution is therefore length biased.
This is why cross-sectional samples of ongoing disease, unemployment, hospitalization or device downtime can overrepresent long-duration cases even when episode onset rates are constant.
The mechanism is not confounding in the ordinary regression sense.
It is selection induced by exposure time.
A duration of ten days creates ten times as much opportunity for cross-sectional observation as a duration of one day.
The same issue appears in status dashboards.
If an operations team opens a dashboard at random times and records the outages currently active, the sample is weighted toward long outages.
The average outage duration among active incidents is not the average outage duration among all incidents.
Both are valid quantities.
They answer different questions.
MTBF is not automatically the expected wait seen at a random time
Mean time between failures is often used as though it described how long an observer should expect to wait until the next failure.
For a renewal failure process with iid operating intervals,
A random-time observer instead sees expected residual operating life
These are equal only under special interval distributions, including the exponential case.
If operating times are nearly deterministic,
then
If operating times are highly variable,
then
The mean alone is therefore insufficient.
Two systems can share the same MTBF and have different random-time residual-life distributions because their interval variances differ.
For maintenance planning, this distinction matters whenever the observation time is not synchronized with a renewal epoch.
A planner commissioning a brand-new component begins at age zero.
A technician inspecting an installed component at an arbitrary calendar time sees a length-biased age distribution.
Those are different prediction problems.
Alternating renewal processes turn the same idea into availability
Suppose a system alternates between an up period
and a down period
The cycle length is
If cycles are iid and have finite means, the renewal-reward theorem gives the long-run fraction of time the system is operational:
This is a time-weighted quantity.
It is not the fraction of episodes that are up, because every cycle contains exactly one up and one down episode.
If down durations vary substantially, a random inspection disproportionately encounters long outages.
The long-run unavailability is
while the duration distribution of the outage seen conditional on observing the system down is the length-biased version of $D$.
Hence even after availability is known, the expected duration of the currently observed outage is
and the expected remaining downtime, under equilibrium sampling, is
A system can have acceptable long-run availability while producing very long residual waits during the relatively rare outages that dominate observed downtime.
This is another reason means alone are not enough for service-level planning.
Event sampling and time sampling are different probability measures
The inspection paradox is often presented as a clever puzzle.
Its deeper mathematical lesson is about sampling measures.
Event-index sampling gives equal weight to each interval.
Time-index sampling gives weight proportional to interval duration.
If
is any function of interval length, its event-average expectation is
The corresponding time-sampled expectation over the enclosing interval is
The factor
is the change of measure.
Setting
gives
Setting
gives
Large intervals are overrepresented in every time-sampled functional.
This perspective extends naturally to Palm probability, which formalizes the difference between observing a stochastic process at a typical time and observing it from a typical event.
A typical event sees one probability law.
A typical time sees another.
Many apparent paradoxes in point processes are really failures to specify which observer is being used.
The bus paradox depends on headway variability, not on buses
Suppose buses have iid headways $X$ and a passenger arrives independently at a random time in equilibrium.
The expected passenger wait is
Write this as
If buses arrive exactly every ten minutes,
and the expected wait is five minutes.
If headways are exponential with mean ten minutes,
and the expected wait is ten minutes.
If bunching produces
then
and the expected wait becomes
minutes.
The paradox is therefore not that a passenger always waits longer than the published mean headway.
The paradox is that a random passenger does not sample headways uniformly.
Variability controls the difference.
The same formula applies to inspection intervals, packet gaps, machine failures, service episodes and many other renewal settings.
The bus is only the mnemonic.
Random observation overweights persistence
The two-point example contains the whole argument.
Ninety percent of event intervals last one hour.
Only ten percent last ten hours.
The ordinary mean is 1.9 hours.
A random time nevertheless lands in a ten-hour interval with probability about 52.6%.
The enclosing interval averages 5.74 hours.
The expected remaining wait is 2.87 hours.
No estimator failed.
No sample was too small.
The observer changed.
Renewal theory makes this distinction precise. Sampling intervals by event index recovers the original inter-arrival distribution. Sampling the process at a random calendar time produces the size-biased law. The equilibrium age and residual life inherit the second moment, which is why variability matters so strongly.
This is the general lesson of the inspection paradox.
Long states are easier to observe because they persist.
Whenever observation probability grows with duration, the observed sample is not a neutral sample of episodes.
The sampling clock has become part of the probability model.
References
Asmussen, S. (2003). Applied Probability and Queues (2nd ed.). Springer.
Cox, D. R. (1962). Renewal Theory. Methuen.
Feller, W. (1971). An Introduction to Probability Theory and Its Applications, Volume II (2nd ed.). Wiley.
Ross, S. M. (2014). Introduction to Probability Models (11th ed.). Academic Press.
Smith, W. L. (1958). Renewal theory and its ramifications. Journal of the Royal Statistical Society: Series B, 20(2), 243–302.
Vardi, Y. (1982). Nonparametric estimation in the presence of length bias. The Annals of Statistics, 10(2), 616–620.
Embed interactive plots, widgets, and demos using <figure>, <iframe>, or <div class="interactive-embed"> containers. Ensure each embed includes descriptive captions for accessibility.
How to cite
Use the quick export buttons to save citations for reference managers or copy the formatted text directly.
Diogo Ribeiro (2026). The Inspection Paradox Is Length-Biased Sampling. Faculty of Media Arts and Design, Technical University of Porto. https://diogoribeiro7.github.io/mathematics/the_inspection_paradox_is_length_biased_sampling/.


