Having had experience of data collection and dissemination, I do wonder about the power of some open source data cleansing here. I was trained on the mantra of "if it is interesting it is wrong", but this is such a big dataset that I doubt that ORR have the resources to identify and check every interesting number.
Two that I have already noticed:
Kings Lynn: the five most popular destinations are Kings Cross, Cambridge, Downham Market, Ely, Watlington, but sixth is, er, Dartford?
(Dartford is also highly ranked at Watlington and Downham Market)
Pulborough: the four most popular destinations are Victoria, Horsham, Gatwick, Chichester, but fifth is, er, Hatfield?
(Hatfield is also highly ranked at Billingshurst and Amberley-I picked these up from looking at Hatfield)
I suspect that there are some coding errors here that have not been cleaned out.
Here the Central/Queen Street numbers look like a possible 60/40 split of Glasgow BR tickets?