mscroggs.co.uk
mscroggs.co.uk

subscribe

Blog

A 20,000-to-1 baby?

 2018-03-23 
This morning, I heard about Arnie Ellis on the Today programme. Arnie is the first baby boy to be born in his family in five generations, following ten girls. According to John Humphrys, there is a 20,000-to-1 chance of this happening. Pretty quickly, I started wondering where this number came from.
After a quick Google, I found that this news story had appeared in many of today's papers, including the Sun and the Daily Mail. They all featured this 20,000-to-1 figure, which according to The Sun originally came from Ladbrokes.

What is the chance of this happening?

If someone is having a child, the probability of it being a girl is 0.5. The probability of it being a boy is also 0.5. So the probaility of having ten girls followed by a boy is
$$\left(\tfrac12\right)^{10}\times\tfrac12=\frac1{2048}.$$
If all 11 children were siblings, then this would be the chance of this happening—and it's a long way off the 20,000-to-1. But in Arnie's case, the situation is different. Luckily in the Daily Mail article, there is an outline of Arnie's family tree.
Here, you can see that the ten girls are spread over five generations. So the question becomes: given a baby, what is the probability that the child is male and his most recently born ten relatives on their mother's side are all female?
Four of the ten relatives are certainly female—Arnie's mother, grandmother, great grandmother and great great grandmother are all definitely female. This only leaves six more relatives, so the probability of a baby being in Arnie's position is
$$\left(\tfrac12\right)^{6}\times\tfrac12=\frac1{128}.$$
This is now an awful lot lower than the 20,000-to-1 we were told. In fact, with around 700,000 births in the UK each year, we'd expect over 5,000 babies to be born in this situation every year. Maybe Arnie's not so rare after all.
This number is based on the assumption that the baby's last ten relatives are spread across five generations. But the probability will be different if the relatives are spread over a different number of generations. Calculating the probability for a baby with any arrangement of ancestors would require knowing the likelihood of each arrangement of relatives, which would require a lot of data that probably doesn't exist. But the actual anwer is probably not too far from 127-to-1.

Where did 20,000-to-1 come from?

This morning, I emailed Ladbrokes to see if they could shed any light on the 20,000-to-1 figure. They haven't got back to me yet. (Although they did accidentally CC me when sending the query on to someone who might know the answer, so I'm hopeful.) I'll update this post with an explanaation if I do hear back.
Until then, there is one possible explanation for the figure: we have looked at the probability that a baby will be in this situation, but we could instead have started at the top of the family tree and looked at the probability that Beryl's next ten decendents were girls followed by a boy. The probability of this happening will be lower, as there is a reasonable chance that Beryl could have no female children, or no children at all. Looking at the problem this way, there are more ways for the situation to not happen, so the probability of it happening is lower.
But working the actually probability out in this way would again require data about how many children are likely in each generation, and would be a complicated calculation. It seems unlikely that this is what Ladbrokes did. Let's hope they shed some light on it...
                        
(Click on one of these icons to react to this blog post)

You might also enjoy...

Comments

Comments in green were written by me. Comments in blue were not written by me.
@Steve Spivey: Nothing
Matthew
                 Reply
Any response from Ladbrokes yet?
Steve Spivey
                 Reply
 Add a Comment 


I will only use your email address to reply to your comment (if a reply is needed).

Allowed HTML tags: <br> <a> <small> <b> <i> <s> <sup> <sub> <u> <spoiler> <ul> <ol> <li> <logo>
To prove you are not a spam bot, please type "nogaced" backwards in the box below (case sensitive):

Archive

Show me a random blog post
 2026 

Feb 2026

Christmas (2025) is over
 2025 
▼ show ▼
 2024 
▼ show ▼
 2023 
▼ show ▼
 2022 
▼ show ▼
 2021 
▼ show ▼
 2020 
▼ show ▼
 2019 
▼ show ▼
 2018 
▼ show ▼
 2017 
▼ show ▼
 2016 
▼ show ▼
 2015 
▼ show ▼
 2014 
▼ show ▼
 2013 
▼ show ▼
 2012 
▼ show ▼

Tags

graph theory interpolation triangles misleading statistics news weak imposition edinburgh raspberry pi coventry crosswords mathslogicbot london underground matrices bots trigonometry asteroids flexagons ucl christmas sobolev spaces dragon curves hyperbolic surfaces reuleaux polygons royal baby pi approximation day kings databet fractals dataset rhombicuboctahedron games polynomials signorini conditions machine learning gather town runge's phenomenon stickers propositional calculus mean game show probability coins youtube matrix multiplication oeis boundary element methods latex gerry anderson pac-man data visualisation curvature regular expressions exponential growth nonograms nine men's morris puzzles folding tube maps cambridge big internet math-off kenilworth hexapawn chalkdust magazine thirteen zines computational complexity countdown finite element method sorting fence posts pascal's triangle error bars light platonic solids tmip statistics harriss spiral crossnumber inline code estimation errors a gamut of games video games newcastle speed dates crossnumbers european cup plastic ratio finite group weather station books advent calendar draughts london alphabets guest posts matrix of cofactors numbers ternary inverse matrices recursion noughts and crosses javascript live stream map projections rust matrix of minors golden spiral sport chess electromagnetic field friendly squares manchester golden ratio accuracy simultaneous equations approximation bempp frobel chebyshev christmas card php fonts mathsjam royal institution numerical analysis menace mathsteroids tennis binary standard deviation pizza cutting manchester science festival datasaurus dozen wool logic matt parker geometry dinosaurs national lottery geogebra pi game of life pythagoras talking maths in public bodmas stirling numbers folding paper realhats arithmetic 24 hour maths partridge puzzle determinants graphs reddit palindromes the aperiodical radio 4 people maths python data squares phd braiding world cup convergence football cross stitch correlation sound warwick wave scattering go bubble bobble programming probability craft gaussian elimination martin gardner hats anscombe's quartet quadrilaterals turtles rugby logo captain scarlet crochet preconditioning hannah fry logs final fantasy

Archive

Show me a random blog post
▼ show ▼
© Matthew Scroggs 2012–2026