Introduction to Random Variables
A random variable (RV) assigns a numerical value to each outcome of a random experiment. Think of it as a rule that converts the abstract outcome (e.g., "heads") into a number you can work with (e.g., ).
Discrete vs. Continuous Random Variables
| Type | Values | Example | Practical business example |
|---|---|---|---|
| Discrete | Countable (often integers); can be finite or countably infinite | Number of heads in 10 coin tosses: | Number of orders from 5 sales calls, number of defective radios in a shipment of 50, number of customers at a restaurant in a day |
| Continuous | Any value in an interval (including fractions) | Time between customer arrivals at a bank: | Actual fluid ounces in a “12 oz” soft‑drink can (between 11.5 and 12.5), temperature at which a chemical reaction occurs (between 100°C and 150°C) |
Note: In the same setting (e.g., a bank), the number of customers arriving is a discrete RV, while the time between arrivals is a continuous RV.
Probability Mass Function (PMF)
For a discrete RV , the probability mass function (PMF), denoted , gives the probability that equals exactly a particular value :
- for all
- (total probability mass = 1)
Why “mass”? Think of 1 unit of probability being distributed across the possible values — the PMF decides how much mass lands on each value.
Example: Fair die
Plot: six bars of equal height ().
Example: Biased die
| 1 | |
| 2 | |
| 3 | |
| 4 | |
| 5 | |
| 6 | |
| Sum | 1 |
Plot: bars of different heights (tallest at ).
Cumulative Distribution Function (CDF) and Tail Probability
The cumulative distribution function (CDF), , gives the probability that is less than or equal to :
The tail probability (complement) is:
Worked example: Fair die
Worked example: Biased die (probabilities as above)
- (Equivalently, )
Expected Value, Variance, and Standard Deviation
Expected value is the probability‑weighted average of all possible values:
- Provides the center of the distribution.
- It need not be a value the RV can actually take (e.g., fair die: , yet you never roll a 3.5).
Variance measures the spread — the probability‑weighted average of squared deviations from the mean:
Standard deviation is the square root of the variance (same units as X):
Worked example: Fair die
Worked example: Biased die
From the PMF table above:
Exam tip: For any discrete RV, always check that the probabilities sum to 1 before computing expected value or variance. Also remember: may be a number the RV never actually takes — it describes the long‑run average, not a possible outcome.
Key takeaways
- A random variable maps random outcomes to numbers; discrete RVs take countable values, continuous RVs take any value in an interval.
- The PMF gives and sums to 1.
- The CDF gives ; tail probability .
- Expected value is the probability‑weighted mean.
- Variance ; standard deviation = .
- Both fair and biased dice illustrate that probabilities change the center and spread, even though the possible values stay the same.
Key Concepts: PMF, CDF, Tail Probability, Expectation, Variance
For a discrete random variable :
- Probability Mass Function (PMF): . Non‑negative and sums to 1 over all .
- Cumulative Distribution Function (CDF): .
- Tail probability: .
- Expected value: .
- Variance: ; standard deviation .
From Data to PMF
When the true PMF is unknown, estimate it using relative frequencies from observed data:
This gives a valid PMF because frequencies are non‑negative and sum to 1.
Example 1: Biased Die (Reviewing PMF, CDF, Tail)
A biased die has PMF:
- Probability of even value: .
- Probability of values between 2 and 5 inclusive: .
- Tail probability : , or .
Approach to any question:
- Translate the question into an event.
- Identify the range of values that satisfy the event.
- Sum the corresponding .
Example 2: Defective Pumps at Kirloskar Factory
Setup: = number of defective pumps per day. Data from 200 days:
| (defects) | Days | Estimated |
|---|---|---|
| 0 | 80 | 0.40 |
| 1 | 50 | 0.25 |
| 2 | 40 | 0.20 |
| 3 | 10 | 0.05 |
| 4 | 20 | 0.10 |
| 0 | 0.00 |
Questions answered using the PMF:
- .
- .
- Expected value:
- Variance and standard deviation:
Exam tip: The steps are identical for any discrete distribution: obtain (here from data), then compute probabilities via summation, expectation as a weighted sum, and variance as the weighted squared deviation.
Example 3: Late Take-offs at Nagpur Airport
Setup: = number of planes taking off late between 9 pm and 11 pm. Data from 30 days:
| (late planes) | Days | Estimated |
|---|---|---|
| 0 | 3 | 0.10 |
| 1 | 6 | 0.20 |
| 2 | 9 | 0.30 |
| 3 | 6 | 0.20 |
| 4 | 6 | 0.20 |
| 0 | 0.00 |
The PMF is derived from the given statements: , , , , , .
Questions answered:
- .
- .
- Expected value:
- Variance and standard deviation:
Key Takeaways
- PMF from data: relative frequencies estimate probabilities; sum to 1.
- Calculating probabilities: add for the event’s range; use for tail probabilities.
- Expectation: weighted average .
- Variance & std dev: weighted average of squared deviations; gives spread in original units.
- Procedure is universal – same steps for any discrete random variable once is known.
Linear Combinations of Random Variables
A linear combination of random variables is a new random variable formed by multiplying each original variable by a constant and summing them: . Why does this matter? Because quantities like total cost, total profit, portfolio returns, and sample averages are all linear combinations. Understanding their expectation and variance lets us predict average outcomes and measure risk.
Linear Transformation of a Single Variable
If (a linear function of a single random variable ), then
- Expectation: Intuition: scaling and shifting the variable scales and shifts its mean.
- Variance: Intuition: shifting () does not affect spread; scaling () multiplies spread by .
Example 1: Temperature conversion Ideal frying temperature in Celsius: with , . Convert to Fahrenheit: .
Example 2: Real estate price (rupees to dollars) Price in rupees: with , . Exchange rate → .
Sum of Two Random Variables
For :
- Expectation: (always additive).
- Variance: where .
If and are independent, then and:
Example: Store sales Store A: , Store B: , Independent.
- Total sales : , , .
- Difference : , , (same variance because ).
Key insight: For independent variables, the variance of a difference is the sum of variances, not the difference.
General Linear Combination
Let . Then:
-
Expectation (always):
-
Variance (if independent):
-
Variance (if not independent): add pairwise covariance terms:
Sum and Difference of Two Independent Identical Variables
If and are independent, , , then:
| Variable | Expectation | Variance |
|---|---|---|
Both have the same variance — the sign of the constant does not affect the squared scaling.
Portfolio of Two Stocks (Correlated)
Shweta Kulkarni invests in Siemens () and Trent (). , ; , ; correlation .
Portfolio return: , where = fraction in Siemens.
Formulas:
For (equal split):
Exploring other splits:
| 0.5 | 4.0% | 1.21% |
| 0.9 | 3.2% | 1.19% |
The 90-10 portfolio is dominated by the 50-50 portfolio: same risk but lower return. After removing dominated choices, the remaining set of portfolios are Pareto optimal — each offers a different risk-return trade-off. Shweta’s optimal split depends on her risk appetite.
Exam tip: Negative correlation reduces portfolio variance. Even without correlation, diversification often lowers risk relative to a single asset.
Sum of Independent Identical Variables
Let be i.i.d. with mean and variance . Define . Then:
Example: Luggage weight Per passenger: kg, kg, independent. kg, kg.
Average of Independent Identical Variables
Define (the sample mean). Then:
Example: Exam scores Per student: , , independent. , , .
Key insight: Averaging reduces variance — the more data points, the more precise the estimate of the mean.
Coefficient of Variation
The coefficient of variation (CV) is a standardized measure of dispersion:
It expresses risk relative to the average — useful for comparing variability across different scales.
- Luggage total: .
- Exam average: .
Key takeaways
- ; .
- For sums: expectations add always; variances add only if independent; with dependence, include covariance.
- Portfolio return variance is reduced by negative correlation between assets.
- The sample mean has variance — a foundation for statistical inference.
- Coefficient of variation standardizes dispersion for comparing risk across different means.
Recap: Discrete Random Variables
A discrete random variable takes only integer values (e.g., 0, 1, 2, …). Its behaviour is fully described by the probability mass function (PMF) . From the PMF we derive:
- Cumulative distribution function (CDF):
- Expectation (mean):
- Variance:
- Standard deviation:
Linear Combinations of Random Variables
For constants and random variables , define . Then:
- Expectation:
- Variance:
The “extra terms” are all pairwise covariances: . If the are independent, all covariances are zero and the variance simplifies to the sum of the squared-coefficient times variances.
Key insight: Linearity of expectation always holds; variance only simplifies under independence.
Key takeaways
- PMF gives probabilities for each outcome; CDF accumulates them.
- is the probability‑weighted average; measures spread around that average.
- For linear combinations: expectation is linear, variance adds covariances (which vanish under independence).
Bernoulli Random Variable
A Bernoulli random variable models an experiment with exactly two outcomes: success () or failure (). Let the success probability be , and the failure probability. The PMF is:
Why it matters
Before studying the Binomial distribution (which counts successes in multiple trials), we must master the Bernoulli – it is the building block. Many real‑world events can be classified as success/failure:
| Context | Success (Y=1) | Failure (Y=0) | |
|---|---|---|---|
| Coin toss (fair) | Heads | Tails | 0.5 |
| Sales call | Customer purchases | No purchase | given |
| Customer survey | Satisfied | Not satisfied | given |
| Daily demand > 10 | Demand > 10 | Demand ≤ 10 | given |
Expectation and Variance
Using the definitions for a discrete random variable:
Exam tip: The expectation of a Bernoulli is simply – the proportion of successes. Its variance is ; memorise these two formulas.
Worked Example: Biased Die
For this biased die, the event “outcome is 5 or 6” is given probability . The individual face probabilities are not needed for the Bernoulli calculation.
- Success: die shows 5 or 6 →
- Failure: die shows 1-4 →
Examples from Earlier Contexts
In each case the Bernoulli is defined by partitioning outcomes of a known discrete distribution:
| Context | Success condition | |||
|---|---|---|---|---|
| Unbiased die | ≤ 2 | |||
| Biased die | ≥ 5 | |||
| Defective pumps (Sangli) | ≤ 1 defective | |||
| Late planes (Sonegaon) | ≤ 2 late | |||
| Insurance call (Priya) | Purchase | |||
| Star day (Baburao) | Sales > ₹10,000 |
How to Create a Bernoulli from Any Discrete Random Variable
Any discrete random variable can be reduced to a Bernoulli by grouping outcomes into two categories. The examples above show exactly this: for the number of defective pumps, “≤1” becomes success and “≥2” failure, preserving the original probabilities.
Key takeaways
- Bernoulli has two outcomes: (success) with , (failure) with .
- , – derived directly from the definitions.
- It is the foundation for the Binomial distribution (multiple independent Bernoulli trials).
Binomial Random Variables and its Distribution
The binomial random variable counts the number of successes in a fixed number of independent Bernoulli trials. Intuitively: if you flip a coin 10 times, how many heads do you get? That count is a binomial random variable.
Binomial Experiment
A binomial experiment has four properties:
- A sequence of identical and independent trials.
- Each trial has exactly two outcomes: success (probability ) and failure (probability ).
- The probability of success is constant across all trials.
- The random variable = number of successes in trials.
| Parameter | Meaning | Example (coin) |
|---|---|---|
| Fixed number of trials | 10 tosses | |
| Success probability per trial | = P(heads) | |
| Observed count of successes | # heads in 10 tosses |
Exam tip:The binomial random variable only counts successes – it ignores the order in which they occur. All arrangements of successes and failures are grouped into the single outcome .
Probability Mass Function (PMF)
The PMF of gives the probability of exactly successes:
where counts the number of ways to arrange successes among trials.
Why this formula works
- : probability of successes.
- : probability of failures.
- : number of distinct sequences containing exactly successes.
Examples
| Experiment | Definition of success | range | ||
|---|---|---|---|---|
| 10 coin tosses | Heads | 10 | 0–10 | |
| 15 tosses of a fair die | 15 | 0–15 | ||
| 10 tosses of a biased die | 10 | 0–10 |
Expectation and Variance
- Expectation:
- Variance:
- Standard deviation:
Intuition: The binomial random variable is the sum of independent Bernoulli() random variables (each with mean and variance ). Hence the mean and variance simply add up.
Worked example (fair die)
For , :
This means over many sets of 15 rolls, you would expect about 5 successes (rolling ≤2), give or take about 1.8.
Key takeaways
- A binomial random variable requires: fixed , independent trials, two outcomes, constant .
- PMF: .
- Mean ; variance .
- The binomial counts successes – it does not track which trials succeeded.
- Common pitfalls: using binomial when trials are not independent (e.g., sampling without replacement) – use hypergeometric instead.
Exam tip:The binomial PMF sums to 1: . This is the binomial theorem in action. Memorise the mean and variance formulas – they appear in nearly every exam problem.
Examples of Binomial Distributions
The binomial distribution models the number of successes in a fixed number of independent trials, each with the same success probability . The examples that follow show how to compute probabilities, expectations, and perform inverse calculations using a spreadsheet.
Key formulas (for binomial random variable )
- PMF:
- Expected value:
- Variance:
- Standard deviation:
Excel: BINOM.DIST function
| Argument | Meaning | Example |
|---|---|---|
y | number of successes | cell reference |
n | number of trials | 10 |
p | probability of success per trial | 1/6 |
cumulative | FALSE → PMF; TRUE → CDF | FALSE gives |
A template table lists from to , then computes (PMF), (CDF), and (complement) for each . The spreadsheet also automatically updates the mean, variance, and standard deviation when or changes.
Example 1: Biased die (10 tosses, success = 5 or more)
- Trials:
- Success: outcome on a die
- Failure: outcome
- Random variable: = number of successes in 10 tosses
Computed parameters
- Mean:
- Variance:
- Standard deviation:
Probability of at least 3 successes
From the spreadsheet, summing the PMF values for gives:
Inverse probability: find such that
Starting from , the probability is . By incrementally increasing in the spreadsheet cell, the desired probability reaches when .
The goal seek feature in Excel automates this.
Exam tip: An inverse probability problem gives you a target probability and asks for the parameter (or sometimes ) that produces it. The spreadsheet approach (trial and error or goal seek) is efficient but on an exam you may need to reason qualitatively or use a calculator.
Example 2: Kedar Apte’s pump defects (14 days, success = day with ≤1 defect)
- Trials: independent days
- Success: day with one or fewer defective pumps
- Failure: day with two or more defects
- Random variable: = number of success days in 14 days
Computed parameters
- Mean:
- Standard deviation:
Probability of at least 10 success days
From the spreadsheet:
| 10 | 0.202 |
| 11 | 0.137 |
| 12 | 0.069 |
| 13 | 0.023 |
| 14 | 0.002 |
Inverse probability: find such that
Current gives . Increase :
Thus a success probability of approximately (between 0.67 and 0.68) yields just over .
How the spreadsheet approach works (diagram)
Key takeaways
- The binomial distribution is fully determined by and ; all probabilities, mean, and variance follow.
- Spreadsheet templates (with
BINOM.DIST) simplify repeated calculations and allow quick parameter sensitivity analysis. - An inverse probability problem fixes a tail probability and solves for (or ) – often by trial and error or goal seek.
- Both examples illustrate the same workflow: define success → set and → compute probabilities → answer questions about expectation, standard deviation, and tail events.
Binomial Distribution – Examples
A binomial random variable counts the number of successes in independent identical Bernoulli trials, each with success probability (fail probability ).
- PMF:
- Mean:
- Variance:
In practice, cumulative probabilities are obtained from a spreadsheet or binomial tables.
Example 1 – Mangesh Nadkarni (Plane Delays)
- Success = “day with ≤2 delays” → , .
- Trials: independent days.
- = # of success days.
| Event | Probability | Calculation |
|---|---|---|
| At least 18 successes | Sum through | |
| At most 10 failures at least 20 successes | Sum through | |
| Between 15 and 25 successes (inclusive) | Sum through |
Example 2 – Priya (Insurance Calls)
- Success = “call leads to purchase” → , .
- Trials: calls per day.
- = # of successes.
Expected number of successes: . Standard deviation: .
| Event | Probability | Interpretation |
|---|---|---|
| At least 6 successes (≥50% conversion) | Very low chance of a “good” day | |
| At most 3 successes | ~80% chance of “bad” day (≤25% conversion) |
Example 3 – Baburao (Star Days)
- Success = “sales >10 000” → , .
- Trials: days.
- = # of star days.
Expected star days: .
| Event | Probability |
|---|---|
| At most 3 star days | |
| Between 3 and 6 star days (inclusive) | (or by subtraction) |
Redefining Success – The Critical Pitfall
The same problem context can require different definitions of “success”. Always re‑identify and for each question.
Example – Kedar Apte (Defective Pumps)
Original success: “≤1 defect per day” → . But new question: “at least 7 days with no defects” → success = “0 defects” → (from the distribution).
- , .
Another question: “at most 3 days with ≥3 defects” → success = “≥3 defects” → .
- , .
Yet another: “at most one defect on all 14 days” → success = “≤1 defect” (), but event is .
- .
Exam tip: Never carry forward a previous without rechecking what “success” means in the current question. The binomial calculation is mechanical; the hard part is mapping the business question to the correct and .
More Complex Examples – Mangesh Nadkarni (Plane Delays Revisited)
Based on historical data for late planes, with probabilities treated as known:
| Question | Success definition | Event | Probability | ||
|---|---|---|---|---|---|
| At least 15 days with ≤1 plane late | ≤1 plane late | 30 | |||
| At most 5 days with exactly 4 planes late | exactly 4 planes late | 10 | |||
| Exactly 15 days with ≤2 planes late | ≤2 planes late | 30 |
The probability of ≥15 successes with is low; the manager is optimistic.
Divide-and-Conquer Procedure
Key takeaways
- Binomial distribution models counts of successes in a fixed number of independent Bernoulli trials with constant .
- The formulas , , are the core tools.
- The most common errors are misidentifying (success definition) and (number of trials). Always start by clarifying these.
- Use complement or subtraction tricks (e.g., ) when helpful, but spreadsheet sums are straightforward.
- Cumulative probabilities (at most, at least, between) are the most frequent exam questions – practice translating business language into probability events.
Recap of Binomial Distribution
A binomial experiment consists of a sequence of identical and independent trials. Each trial has exactly two outcomes — success (probability ) and failure (probability ). The probability remains constant across trials. The binomial random variable counts the number of successes in such trials. Equivalently, it is the sum of identical and independently distributed (IID) Bernoulli random variables.
If any assumption fails — e.g., success probabilities change or trials are dependent — the experiment is no longer binomial. In practice, other discrete distributions are then used to model the situation.
Key takeaways
- Binomial: fixed , constant , independence, two outcomes per trial.
- with PMF .
- , .
- When assumptions break, alternative distributions are needed.
Poisson Distribution
The Poisson random variable counts the number of occurrences of an event over a specified interval of time or space. It is a discrete distribution (values ) and is often used to model random arrivals, defects, or counts that occur at a constant average rate.
Intuition and Examples
- Number of customers arriving at a store in 30 minutes.
- Number of phone calls at a call center in 15 minutes.
- Number of defects in one kilometre of highway.
- Number of keyword occurrences on a page of text.
Probability Mass Function (PMF)
Let be Poisson with parameter (the mean number of events in the interval). The probability of exactly events is:
is the base of the natural logarithm.
Exam tip: Poisson has only one parameter , unlike binomial’s two ().
Expectation and Variance
For a Poisson random variable :
Thus the mean equals the variance — a unique property of the Poisson distribution. The standard deviation is .
Properties of a Poisson Process
A Poisson process is a stochastic process where events occur randomly over time, and the number of events in any interval of length follows a Poisson distribution. The process satisfies:
- Stationarity – The probability of a given number of events in an interval depends only on the interval’s length, not its start time.
- Independence – The occurrence or non‑occurrence of events in disjoint intervals are independent.
- Mean equals variance – The number of events in any interval has identical mean and variance.
The time between successive events in a Poisson process follows an exponential distribution (a continuous distribution covered later).
Parameter as a Rate
When interest lies in the number of events over an interval of length , the Poisson parameter becomes , where is the rate (events per unit time). Then – the number of events in time – is Poisson with mean :
Worked Example: Call Centre Arrivals
Suppose customers call a help desk at a rate of per hour.
- In a ‑hour period, the number of calls is Poisson with . , .
- In a ‑minute ( hour) period, the number of calls is Poisson with . , .
The PMF in each case uses the corresponding .
Connection to the Mathematician
The distribution is named after Baron Siméon Denis Poisson (1781–1840), a French mathematician and physicist.
Key takeaways
- Poisson models counts of rare events over time or space.
- PMF: , with single parameter .
- .
- For a process with rate over interval , parameter .
- Key properties: stationarity, independence, mean=variance.
Poisson Distribution: Examples and Excel Implementation
The Poisson distribution models the number of events occurring in a fixed interval of time (or space) when events happen at a constant average rate and independently. The distribution has a single parameter (also denoted as the mean and variance). The probability mass function (PMF) is:
Setting Up Poisson in Excel
Excel does not have a dedicated Poisson function, but the PMF can be built using standard functions:
| Cell | Content | Purpose |
|---|---|---|
| D3 | (mean) | Input parameter |
| E3 | =D3 (variance) | Since |
| F3 | =SQRT(E3) | Standard deviation |
| B6:B106 | Values of the random variable | |
| C6 | =EXP(-D3) * D3^B6 / FACT(B6) | PMF: |
| D6 | =C6 | CDF for (since ) |
| D7 | =C7 + D6 (drag down) | CDF: |
| E6 | =1 | Tail probability for : |
| E7 | =E6 - C6 | Tail: (drag down for ) |
This table allows quick computation and plotting of the PMF for any .
Example 1: Customer Arrivals at an Eatery
At Sheetal Dadava, the number of customers arriving in one hour follows a Poisson distribution with . Questions and solutions:
| Question | Probability Statement | Excel Lookup | Result | |
|---|---|---|---|---|
| Exactly 8 customers in 1 hour | 10 | Column C, | 0.113 | |
| More than 10 customers in 1 hour | 10 | Column E, | 0.417 | |
| Exactly 4 customers in 30 minutes | 5* | Column C, | 0.175 | |
| At most 10 customers in 30 minutes | 5* | Column D, | 0.986 |
* scales with time: for a 30‑minute interval, (the process is Poisson; the rate is proportional to interval length).
Exam tip: When the time interval changes, always rescale . “More than 10” means , which equals . Reading the wrong row is a common trap.
Example 2: Social Media Likes
Sriram’s weekly likes follow a Poisson distribution with . Similarly, daily likes use .
| Question | Probability Statement | Excel Lookup | Result | |
|---|---|---|---|---|
| At least 10 likes in a week | 14 | Column E, | 0.891 | |
| Between 10 and 20 likes in a week (inclusive) | 14 | Sum of column C for to | 0.843 | |
| At least 5 likes on a given day | 2 | Column E, | 0.053 | |
| No likes on a given day | 2 | Column C, | 0.135 |
Exam tip: For “between … and …” always check inclusiveness. If both endpoints are included, sum the PMF values directly. The tail probability column gives , not .
Key Takeaways – Poisson Examples
- The Poisson PMF is ; = mean = variance.
- Excel setup (no built‑in function) uses
EXP,^, andFACTto compute PMF, then cumulative and tail probabilities by iteration. - When the time interval changes, scales proportionally (e.g., half the time → half the ).
- Translating a business question into a probability statement is critical: “more than 10” = ; “at most 10” = ; “at least 10” = .
- For range probabilities (between a and b inclusive), sum the PMF for each in that range.
Examples of Poisson Distribution (Call Center)
The Poisson distribution models the number of events (e.g., phone calls) occurring in a fixed interval of time when events happen independently at a constant average rate. Its only parameter is the mean (the average number of events per interval). The probability mass function is
A classic application is call‑center staffing. A call center receives calls according to a Poisson process: the number of calls in any interval depends only on the interval length, not on the time of day, and successive intervals are independent. (More advanced models use non‑stationary Poisson processes with time‑varying rates.)
Scaling the Poisson Parameter
The average rate is given for one interval but may be needed for another. If the rate is constant, scales linearly with the interval length.
Example: The call center u.net averages 6 calls per 15 minutes. Rate per minute = calls/min. For any ‑minute interval, .
| Interval | calculation | Query | Result (from PMF) | |
|---|---|---|---|---|
| 5 minutes | 2 | |||
| 15 minutes | (given) | 6 | ||
| 3 minutes | 1.2 |
Worked Calculations
1. Probability of 3 calls in 5 minutes
, :
There is roughly an 18% chance of exactly 3 calls in a 5‑minute window.
2. Probability of 10 calls in 15 minutes
, :
A 4.1% chance – a relatively low probability, suggesting such high demand may strain staffing.
3. Probability of no calls during a 3‑minute break
, :
Only a 30% chance that a 3‑minute break will be uninterrupted. The agent is likely to be called back before the break ends.
Key Takeaways
- The Poisson distribution depends entirely on the mean for the given interval.
- Scaling: for a new interval = (original rate per unit time) × (new interval length).
- Probabilities are computed directly from the PMF (or a spreadsheet).
- Call‑center applications illustrate how randomness affects service quality and staffing decisions.
Exam tip: Always check that the you use matches the interval of the question. A common mistake is to use the original without rescaling.
Poisson Distribution – Worked Examples
The Poisson distribution models the count of events occurring in a fixed interval of time or space when events happen independently at a constant average rate. The single parameter (often called the rate parameter) equals both the mean and variance:
When the interval changes, scales proportionally.
Example 1: Mall Visitor Arrivals
Context: Prozone Mall, Aurangabad. Average visits during busy hour = 180 per hour. Manager Mansoor wants the probability of 10–20 people accumulating in a 10‑minute window.
Adjusting for the new interval
- 180 per hour 10 minutes: visitors per 10 min
- Similarly, if average is 150/hr → ; if 210/hr →
Probabilities (from Poisson PMF)
| Average per hour | (10 min) | |
|---|---|---|
| 150 | 25 | 0.73 |
| 180 | 30 | 0.526 |
| 210 | 35 | 0.225 |
Why does the probability decrease as increases? The Poisson distribution becomes roughly symmetric around its mean. As shifts right (25 → 30 → 35), the fixed window [10,20] captures less probability mass because the distribution moves to higher counts.
Exam tip: Always re‑scale to match the interval of interest. The probability over the same numeric range can rise or fall as changes – your intuition about “more arrivals = more in the window” fails when the window is far from the mean.
Example 2: Highway Defects
Context: Maharashtra State Highway No. 3. Average defects = 3 per km. Count of defects in a stretch follows Poisson.
Probabilities for 1 km stretch ()
- (tail probability)
- (cumulative)
Probabilities for an 8 km stretch ()
Common mistake: Trying to compute as (or as initially attempted). This is wrong because it assumes every 1‑km segment must have ≤5 defects – the total can be ≤40 even if some segments exceed 5, as long as others compensate.
Exam tip: For a Poisson process, the total count over combined independent intervals is Poisson with . NEVER multiply probabilities of sub‑intervals.
Example 3: Social Media Flags
Context: NIA analyst Nitin monitors flags per conversation. Average = 5 flags per conversation. Flags occur independently and at constant rate.
Single conversation ()
- (less than 3% – low alert)
Two consecutive conversations ()
- (very unlikely)
Note that “less than 4 flags in two conversations” means , not .
Result summary table
| Scenario | Query | Probability | |
|---|---|---|---|
| 1 conversation | 5 | 0.891 | |
| 1 conversation | 5 | 0.032 | |
| 2 conversations | 10 | 0.963 | |
| 2 conversations | 10 | 0.01 |
Key Properties of the Poisson Distribution
- Constant rate – probability of an event is the same in any two intervals of equal length.
- Independence – occurrence in one interval is independent of occurrence in any other non‑overlapping interval.
- Mean = Variance = .
These properties make Poisson suitable for counts over time (arrivals, calls, flags) and over space (defects per km, potholes, errors in wafers or text).
Connection Between Binomial and Poisson
When (success probability) is small and is large, the binomial distribution approximates a Poisson with .
- Binomial: ,
- Poisson:
If is small, , so binomial variance ≈ , matching Poisson. Many Poisson distributions also appear symmetric and bell‑shaped for moderate – a preview of the Central Limit Theorem.
Exam tip: Use the Poisson approximation to the binomial when is large, is small, and is moderate (rule of thumb: , ).
Key Takeaways
- Poisson models counts of rare, independent events over time/space; one parameter = mean = variance.
- When interval changes, scales proportionally; always recompute for the exact interval.
- Probabilities computed via PMF, cumulative (CDF), or tail – spreadsheet tools are helpful.
- Common pitfalls: multiplying probabilities across intervals instead of summing means; confusing with .
- The binomial approximates Poisson when is small ().