Skip to main content

Everything is Predictable - Tom Chivers *****

There's a stereotype of computer users: Mac users are creative and cool, while PC users are businesslike and unimaginative. Less well-known is that the world of statistics has an equivalent division. Bayesians are the Mac users of the stats world, where frequentists are the PC people. This book sets out to show why Bayesians are not just cool, but also mostly right.

Tom Chivers does an excellent job of giving us some historical background, then dives into two key aspects of the use of statistics. These are in science, where the standard approach is frequentist and Bayes only creeps into a few specific applications, such as the accuracy of medical tests, and in decision theory where Bayes is dominant.

If this all sounds very dry and unexciting, it's quite the reverse. I admit, I love probability and statistics, and I am something of a closet Bayesian*), but Chivers' light and entertaining style means that what could have been the mathematical equivalent of debating angels on the heads of a pin becomes both enthralling and relatively easy to understand. You may have to re-read a few sentences, because there is a bit of a head-scrambling concept at the heart of the debate - but it's well worth it.

A trivial way of representing the difference between Bayesian and frequentist statistics is how you respond to the question 'What's the chance of the result being a head?' when looking at a coin that has already been tossed, but that you haven't seen. Bayesian statistics takes into account what you already know. As you don't know what the outcome is, you can only realistically say it's 50:50, or 0.5 in the usual mathematical representation. By contrast, frequentist statistics says that as the coin has been tossed, it is definitely heads or tails with probability 1... but we can't say which. This seems perhaps unimportant - but the distinction becomes crucial when considering the outcome of scientific studies.

Thankfully, Chivers goes into in significant detail the problem that arises because in most scientific use of (frequentist) probability, what the results show is not what we actually want to know. In the social sciences, a marker for a result being 'significant' is a p-value of less that 0.05. This means that if the null hypothesis is true (the effect you are considering doesn't exist), then you would only get this result 1 in 20 times or less. But what we really want to know is not the chance of this result if the hypothesis is true, but rather what's the chance that the hypothesis is true - and that's a totally different thing.

Chivers gives the example of 'it's the difference between "There's only a 1 in 8 billion chance that a given human is the Pope" and "There's only a 1 in 8 billion chance that the Pope is human"'. At risk of repetition because it's so important, frequentist statistics, as used by most scientists, tells us the chance of getting the result if the hypothesis is true; Bayesian statistics works out what the chance is of the hypothesis being true - which most would say is what we really want to know. In fact, as Chivers points out, most scientists don't even know that they aren't showing the chance of the hypothesis being true - and this even true of many textbooks for scientists on how to use statistics.

At this point, most normal humans would say 'Why don't those stupid scientists use Bayes?' But there is a catch. To be able to find how likely the hypothesis is, we need a 'prior probability' - a starting point which Bayes' theorem then modifies using the evidence we have. This feels subjective, and for the first attempt at a study it certainly can be. But, as Chivers points out, in many scientific studies there is existing evidence to provide that starting point - the frequentist approach throws away this useful knowledge.

Is the book perfect? Well, I suspect as a goodish Bayesian I can never say something is perfect. I found it hard to engage with an overlong chapter called 'the Bayesian brain' that is not about using Bayes, but rather trying to show that our brains take this approach, which all felt a bit too hypothetical for me. And Chivers repeats the oft-seen attack on poor old Fred Hoyle, taking his comment about evolution and 'a whirlwind passing through a junkyard creating a Boeing 747' in a way that oversimplifies Hoyle's original meaning. But these are trivial concerns.

I can't remember when I last enjoyed a popular maths book so much. It's a delight.

* Not entirely a closet Bayesian - my book Dice World includes an experiment using Bayesian statistics to work out what kind of dog I have, given a mug that's on my desk.

Hardback:   
Kindle 
Using these links earns us commission at no cost to you
Review by Brian Clegg - See all Brian's online articles or subscribe to a weekly email free here

Comments

Popular posts from this blog

Mathematics with Love – Mary Stopes-Roe *****

Admittedly it’s early days (this review is written in January), but this, for me, is the surprise hit of the year so far! I approached this book with trepidation, but found it absolutely delightful. It is described on the cover as the “courtship correspondence of Barnes Wallis, inventor of the bouncing bomb”, and contains a series of letters between Wallis and his cousin and eventual wife Molly Bloxham, along with some useful annotation by their daughter, Mary. The courtship itself is not without difficulties, as Wallis was 18 years older than the 17-year-old Molly at the start of the correspondence, and her father, not surprisingly, wasn’t too pleased about the interest of such an elderly suitor, but that isn’t the only reason the letters are interesting – it’s also because of maths, and Wallis’s position in the UK as the engineering hero of the Second World War. (Incidentally, it seemed very strange to see letters addressed to “Barnes” – I had always assumed Barnes Wallis was a ...

Data Empire - Roopika Risam ****

The central thesis presented by Roopika Risam is that information gives us (and particularly countries) the power to organise, control and dominate others. Although I have a couple of issues with the presentation, this is a genuinely interesting trip through the history of our use of stored information from the earliest tallies to the latest information technology. I loved a quote from Lisa Gitelman that data is is always 'cooked' so 'raw data is an oxymoron'. This neatly underlines Risam's thesis that data and information are not neutral facts, but rather tools that (like everything from fire to electronics) can be used for good or evil. As we are taken through the historical context, it can sometimes be a little difficult to judge whether Risam regards a particular example as bad or good, even when the outcome is disastrous. One thing I didn't like too much is the adherence to a popular science writing approach that has got distinctly hackneyed: opening chapte...

Andrew Jaffe - Five way interview

Andrew Jaffe is professor of astrophysics and cosmology at Imperial College, London and director of the Imperial Centre for Inference and Cosmology. His new book is The Random Universe. Why science? I’ve always been interested in science, in particular in  astronomy, astrophysics, and space. One of my earliest memories - back in nursery school in New Jersey, I think - was watching one of the moon launches. I wanted that excitement to be part of my life! I never got to be an astronaut, but I did get to be part of the Planck Satellite team, and was privileged to be able to travel to the ESA Spaceport in French Guiana to watch the launch.  In between, I was lucky enough to have a supportive family, get a good education, and find inspiring teachers, mentors, and collaborators. They helped me model the universe, and helped me learn how to refine those models in the face of experimental and observational evidence. That is, they taught me to be a scientist. Why this book? The Random...