Skip to main content

Causal Inference - Paul Rosenbaum ***

The whole business of how we can use statistics to decide if something is caused by something else is crucially important to science, whether it's about the impact of a vaccine or deciding whether or not a spray of particles in the Large Hadron Collider has been caused by the decay of a Higgs boson. 'Correlation is not causality' is a mantra of science, because it's so easy to misinterpret a causal link from things that happen close together and space and time. As a result I was delighted with the idea of what the cover describes as a 'nontechnical guide to the basic ideas of modern causal inference'.

Paul Rosenbaum starts with a driving factor - deducing the effects of medical treatments - and goes on to bring in the significance of randomised experiments versus the problems of purely observational studies, digs into covariates and ways to bring in experiment-like features to observational studies, brings up issues of replication and finishes with the impact of uncertainty and complexity. This is mostly exactly the kind of topics than should be covered in such a guide, and as such it hits spot. But, unfortunately, while it is indeed an effective introductory guide for scientists who aren't mathematicians, Rosenbaum fails on making this accessible to a nontechnical audience.

Rosenbaum quotes mathematician George Pólya as saying that we need a notation that is 'unambiguous, pregnant, easy to remember…' I would have been happier with this book if Rosenbaum had explained how a mathematical notation could possibly be pregnant. (He doesn't.) But, more importantly, the notation used is simply not easy to remember for a nontechnical audience. Within one page of it starting to be used, I had to keep looking back to see what the different parts meant. 

We are told that a causal effect is 'a comparison of outcomes' and in the first example given this is rTw - rCw. Bits of this are relatively clear. T and C are treatment and control. W is George Washington (as the example is about his being treated, then dying soon after). I'm guessing 'r' refers to result, though that term isn't used in the text, but most importantly it's not obvious why the 'causal effect' is those two variables, set to arbitrary values, with one subtracted from the other. I'm pretty familiar with algebra and statistics, but I rapidly found the symbolic representations used hard to follow - there has to be a better way if you are writing for a general audience: it appears the author doesn't know how to do this. 

The irritating thing is that Rosenbaum doesn't then make use of this representation - he's lost half the readership for no reason. The rest of the book is more descriptive, but time after time the way that examples are described is handled in a way that is going to put people off, bringing in unnecessary jargon and simply writing more like a textbook without detail. Take the opening of the jauntily headed section 'Matching for Covariates as a Method of Adjustment': 'In figure 4 [which is several pages back in a different chapter], we saw more extensive peridontal disease amongst smokers, but we were not convinced that we were witnessing an effect caused by smoking. The figure compared the peridontal disease outcomes of treated individuals and controls who were not comparable. In figures 2-3 we saw that the smokers and nonsmokers were not comparable. The simplest solution is to compare individuals who are comparable, or at least comparable in ways we can see.' 

This is a classic example of the importance of being aware of who the audience is and what the book is supposed to do. To reach that target nontechnical audience, the book would have to have been far less of a textbook light, rethinking the way the material is put across. The content is fine for a technical audience who aren't mathematicians - so this is still a useful book - but the content certainly isn't well-presented for the general public.

Paperback:   
Kindle 
Using these links earns us commission at no cost to you
Review by Brian Clegg - See all Brian's online articles or subscribe to a weekly email free here

Comments

Popular posts from this blog

Mathematics with Love – Mary Stopes-Roe *****

Admittedly it’s early days (this review is written in January), but this, for me, is the surprise hit of the year so far! I approached this book with trepidation, but found it absolutely delightful. It is described on the cover as the “courtship correspondence of Barnes Wallis, inventor of the bouncing bomb”, and contains a series of letters between Wallis and his cousin and eventual wife Molly Bloxham, along with some useful annotation by their daughter, Mary. The courtship itself is not without difficulties, as Wallis was 18 years older than the 17-year-old Molly at the start of the correspondence, and her father, not surprisingly, wasn’t too pleased about the interest of such an elderly suitor, but that isn’t the only reason the letters are interesting – it’s also because of maths, and Wallis’s position in the UK as the engineering hero of the Second World War. (Incidentally, it seemed very strange to see letters addressed to “Barnes” – I had always assumed Barnes Wallis was a ...

Data Empire - Roopika Risam ****

The central thesis presented by Roopika Risam is that information gives us (and particularly countries) the power to organise, control and dominate others. Although I have a couple of issues with the presentation, this is a genuinely interesting trip through the history of our use of stored information from the earliest tallies to the latest information technology. I loved a quote from Lisa Gitelman that data is is always 'cooked' so 'raw data is an oxymoron'. This neatly underlines Risam's thesis that data and information are not neutral facts, but rather tools that (like everything from fire to electronics) can be used for good or evil. As we are taken through the historical context, it can sometimes be a little difficult to judge whether Risam regards a particular example as bad or good, even when the outcome is disastrous. One thing I didn't like too much is the adherence to a popular science writing approach that has got distinctly hackneyed: opening chapte...

Andrew Jaffe - Five way interview

Andrew Jaffe is professor of astrophysics and cosmology at Imperial College, London and director of the Imperial Centre for Inference and Cosmology. His new book is The Random Universe. Why science? I’ve always been interested in science, in particular in  astronomy, astrophysics, and space. One of my earliest memories - back in nursery school in New Jersey, I think - was watching one of the moon launches. I wanted that excitement to be part of my life! I never got to be an astronaut, but I did get to be part of the Planck Satellite team, and was privileged to be able to travel to the ESA Spaceport in French Guiana to watch the launch.  In between, I was lucky enough to have a supportive family, get a good education, and find inspiring teachers, mentors, and collaborators. They helped me model the universe, and helped me learn how to refine those models in the face of experimental and observational evidence. That is, they taught me to be a scientist. Why this book? The Random...