Showing posts with label mapping. Show all posts
Showing posts with label mapping. Show all posts

Monday, 27 April 2015

Voices of Authority: Towards a history from below in patchwork



This post is intended to very briefly describe a project I am about halfway through - that seeks to experiment with the new permeability that digital technologies seem to make possible - to create a more usable 'history from below', made up of lives knowable only through small fragments of information.

This particular project is called ‘Voices of Authority’, and is a small part of a larger AHRC funded project – The Digital Panopticon – that is seeking to digitise and link up the records reflecting everyone tried in London between around 1780 and 1875, and either sent to prison, or else transported to Australia.  This small element of the wider project is bringing together a series of different ways of knowing about a particular place, time and experience – the Old Bailey courtroom from around 1750-1850, and the experience of being tried for your life and for your liberty.  The conceit behind this project, is really a suggestion that building something in three dimensions, with space, physical form and performance, along with new forms of analysis of text; can change how we understand the experience of the trial process; and to allow a more fully empathetic engagement with defendants; along with a better understanding of how their experience impacted on the exercise of power and authority.

This project is only half completed – so this is very much a report of 'work in progress'.  But, in essence, what seeks to do is bring together three distinct different forms of ‘data’ and to re-organise that data around individual defendants.  

First, it takes the text of the Old Bailey Proceedings – the trial accounts of some 197,745 trials held between 1674 and 1913, and recognises them as comprising two different and distinct things – a bureaucratic record of the trials themselves (names, verdicts, punishments); and at the same time, one of the largest corpora of recorded spoken language created prior to the twentieth century – some 40 million words of direct, recorded testimony for the period under analysis.

These understandings of the Proceedings are, of course, built on projects of much longer duration; including the OldBailey Online, and more particularly, on Magnus Huber’s additional linguistic mark-up of the Proceedings, which allows ‘speech’ to be pulled from the trial text, and to identify the speaker along the way.  This is available via the Old Bailey Corpus.



The project also builds on text and data mining methodologies – including direct counting of word and phrase distributions, and the application of a form of explicit semantic analysis, that allows us to look at the changing character of language used in witness statements over the course of the eighteenth and nineteenth centuries.
 

In other words, the first element of the project is the text and speech, crimes, punishments and dates provided by the Old Bailey Proceedings.

The second element is the body of the criminal - the physical body of the individual  men and women involved.  The broader project is creating a dataset of some 66,000 men and women – with substantial and detailed information about their lives, both before and after transportation or imprisonment – reflecting the inter-relationship between the people who became defendants and criminals with the systems of a global empire. And this material provides a huge amount of data about bodies – to add to the words individuals spoke to power.  Height, weight, eye colour, tattoos among a range of other aspects of a physical self.  Suddenly, we know if a collection of words was spoken by a ten stone, 5 foot two inch woman with brown hair and black eyes, and a withered left arm; or by a six foot man with an anchor tattoo on his left arm, and a scar above squinting blue eyes. 

To think about it another way, this bit of the ‘project’, allows us to worry about the ragged boundary between the ‘physical’ as recorded in a set of numerical and standardised descriptions, and the ‘textual’ – the slippery and ambiguous content of each witness statement. 

In relation to ‘history from below’, this allows us to put together the lives of people like William Curtis, who as a 16 year old, in the summer of 1843, had a perfectly healthy tooth pulled, before stealing the dentists’ coat.  And Sarah Durrant, who was convicted for receiving two banks notes worth 2000 pounds. It all allows us to know their words (their textualities), and at the same time to see them as part of a different kind of truth – of place and body. 

The aspiration is to essentially code for the variabililties of body type at scale, to add a further dimension to both the records of the bureaucracy of trials (charge, verdict punishment), and the measurable content of the textualities of those same trials.

And finally, we are adding one additional dimension – space –  a ‘scene of trial’.

For this we are first of all building on a project called Locating London’s Past, which among other things, maps crime locations on to the historic landscape of London.  And to this we are adding a reconstruction of the courtroom, where all these trials took place. 

Simply using Sketch-up, we have made most progress on the George Dance’s building, finished in the late 1770s and providing the main venue for the relevant trials for the next hundred years – basing the models on the architectural plans from which it was built.  



In the process of creating this model, huge amounts of imformation about trial procedure has been revealed, including the changing layout of the court, and the relative position of the different speakers.  The design itself reflects a hitherto unacknowledged transition in the character of how witnesses and defendants were divided in this evolving space, evidencing a new story of the evolution of the criminal trial. 



The architecture itself, suggests that there was a clear transition from a situation in which witnesses and victims stood in a similar relation to the judge and jury (both facing the judge, relatively close to one another); to one - like a modern anglo-american courtroom  - where the judge and witnesses are on one side, and the defendant on the opposite side of the courtroom.  In other words, the character of the adversarial relationship at the core of the adversarial trial was re-defined, with the witnesses and victim re-located on the either side of the argument, and the judges role, redefined as arbiter between them.  At some level, in the process community resolution was replaced by court judgement. 



If you want to explain why conviction rates at the Old Bailey rose from under 50% in the mid eighteenth century, to over eighty percent at the end of the nineteenth century, starting from precisely the moment when the courtroom was rebuilt, this ‘fact on the ground’ needs to be part of the story.

What has also been revealed is the importance of levels – with lawyers speaking upwards to the judge, jury, witnesses and defendant, from a cock-pit several feet below their eye level.  Like a theatre audience, the judge, jury and defendant looked down on the stage below.  In other words, what was created, at least for a short time (70 years), was the real feel of a ‘theatre’ in which, as a barrister, you were forced to perform to the gods.  


 

Looking forwarding to the next stage, this particular sub-project is seeking to move from the ‘art’ of making and performance, through a humanist and historical appreciation of ‘experience’, towards the tools of social science and informatics – seeking to combine the close reading of a single desperate plea, with the empathy that can only come with physical knowledge, with that macroscopic image of all the similar words spoken over a hundred years – how that one plea fits in a universe of words and bodies.  And all of this, is in turn, being undertaken in pursuit of a more nuanced and empathetic engagement with the lives of working people - both for its own sake, and as part of a new analysis of the workings of power 'from below'.

With luck, this will allow us to move beyond a simple analysis of the courtroom, and the ‘adversarial trial’ - to an analysis through which we can see the whole system from the defendants’ perspectiv.

In other words, the next step is about creating a history of the British criminal justice system, and of transportation from an experiential perspective on a large scale – contributing to a history of common human experience, evidenced from the distributed leavings of the dead, analysed with all the approaches available to hand, from all the perspectives available.

Sunday, 9 November 2014

Big Data, Small Data and Meaning

This post was originally written as the text for a talk I gave at a British Library Lab's event in London in early November 2014. In the nature of these things, the rhythms of speech and the verbal ticks of public speaking remain in the prose. It has been lightly edited, but the point remains the same.  

In recent months there has been a lot of talk about big stuff. Between 'Big Data' and calls for a return to ‘Longue durée’ history writing, lots of people seem to be trying to carve out their own small bit of 'big data'. This post represents a reflection on what feels to me to be an important emerging strategy for information interrogation driven by the arrival of 'big data' (a 'macroscope'); and a tentative step beyond that, to ask what is lost by focusing exclusively on the very large. 

And the place I need to start is with the emergence of what feels to me like an increasingly commonplace label – a ‘macroscope’ - for a core aspiration of a lot of people working in the Digital Humanities. 

As far as I can tell, the term ‘macroscope’ was coined in 1969 by Piers Jacob, and used as the title of his science fiction/fantasy work of the same year – in which the ‘macroscope’, a large crystal, able to focus on any location in space-time with profound clarity, is used to produce something like a telescope of infinite resolution. In other words, a way of viewing the world that encompasses both the minuscule, and the massive. The term was also taken up by Joel de Rosnay and deployed as the title of a provocative book on systems analysis first published in 1979. The label has also had a long and undistinguished afterlife as the trademark for a suite of project management tools – a ‘methodology suite’ - supported by the Fujistu Corporation. 

But I think the starting point for interest in the possibility of creating a ‘macroscope’ for the Digital Humanities, comes out of computer science, and the work of Katy Börner from around 2011.
Her designs and advocacy for the development of a ‘Plug and Play Macroscope’, seems to have popularised the idea to a wider group of Digital Humanists and developers. To quote Börner

'Macroscopes provide a "vision of the whole," helping us "synthesize" the related elements and detect patterns, trends, and outliers while granting access to myriad details. Rather than make things larger or smaller, macroscopes let us observe what is at once too great, slow, or complex for the human eye and mind to notice and comprehend.' (Katy Börner, ‘Plug-and-Play Macroscopes’, Communications of the ACM, Vol. 54 No. 3, Pages 60-6910.1145/1897852.1897871)

In other words, for Börner, a macroscope is a visualisation tool that allows a single data point, to be both visualised at scale in the context of a billion other data points, and drilled down to its smallest compass. This was not a vision or project initially developed in the humanities. Instead it was a response to the conundrums of ‘Big Data’ in both STEM academic disciplines, and the wider commercial world of information management. But more recently, a series of ‘macroscope’ projects have begun to emerge from within the humanities, tied to their own intellectual agendas, and subtly recreating the idea with a series of distinct emphases. 

Perhaps the project most heavily promoted recently, is Paper Machines, created by Jo Guldi and Chris Johnson-Robertson – and the MetLab at Harvard. This forms a series of visualisation tools, built to work with Zotero, and ideally allowing the user to both curate a large scale collection of works, and explore its characteristics through space, time and word usage. In other words, it is designed to allow you to build your own Google Books, and explore. There are problems with Paper Machines, and most people I know have struggled to make it work consistently. But it rather nicely builds on the back of functionality made available through Zotero, and effectively illustrates what might be described as a tool for ‘distant reading’ that encompasses elements of a ‘macroscope’. 

What is most interesting about it, however, is the use its creators make of it in seeking to shift a wider humanist discussion from one scale of enquiry to another. Last month, to great fanfare, CUP published Jo Guldi and David Armitage’s History Manifesto, which argues that once armed with a ‘macroscrope’ – Paper Machines in their estimation historians should pursue an analysis of how ‘big data’ might be used to re-negotiate the role of the historian – and the humanities more generally. Basically, what Guldi and Armitage are calling for through both the Manifesto and through Paper Machines, is the re-invention of ‘Longue durée’ history – telling ever larger narratives about grand sweeps of historical change, encompassing millennia of human experience. And to do this in pursuit of taking on the mantle of a public intellectual, able to speak with greater authority to ‘power’. 

In the process they explicitly denigrate notions of ‘micro-history’ as essentially irrelevant. At one and the same time, they seem to me to celebrate the possibility of creating a ‘macroscope’, while abjuring half its purpose. What we see in this particular version of a ‘macroscope’ is a tool that privileges only one setting on the scale between a single data point, and the sum of the largest data set we can encompass. In other words, by seeking the biggest of big stories, it is missing the rest. 

Perhaps the other most eloquent advocate for a ‘macroscope’ at the minute is Scott Weingart. With Shawn Graham and Ian Milligan, he is writing a collective online ‘book’ entitled, Big Digital History: Exploring Big Data through a Historian’s Macroscope. The book is a nice run through of digital humanist tools, but the important text from my perspective is a blog post Weingart published on the 14 September 2014. The post was called: The moral role of DH in a data-driven world; and in it, Weingart advocates a very specific vision of a ‘macroscope’, in which the largest scale of reference and view is made intelligible through the application of a formal version of network analysis. 

Weingart is a convincing advocate for network analysis, performed in light of some serious and sophisticated automated measures of distance and direction. And his work is a long way ahead of much of the naïve and unconvincing use of network visualisations current in large parts of the Digital Humanities. Weingart also makes a powerful case for where a limited number of DH tools – primarily network analysis and topic modelling - could be deployed in re-engaging the ‘humanities’ with a broader social discussion. 

Again, like Guldi and Armitage, Weingart seeks in 'Big Data' a means through which the Humanities can ‘speak to power’. As with the work of Armitage and Guldi, the pressing need to turn Digital Humanities to political account appears to motivate a search for large scale results that can be deployed in competition with the powerful voices of a positivist STEM tradition. My sense is that Weingart, Armitage and Guldi are all essentially scanning the current range of digital tools, and selectively emphasising those that feel familiar from the more ‘Social Science’ end of the Humanities. And that having located a few of them, they are advocating we adopt them in order to secure our place at the table. 

In other words, there is a cultural/political negotiation going on in these developments and projects that is driven by a laudable desire for ‘relevance’, but which effectively moves the Humanities in the direction of a more formal variety of Social Science. 

Others still, are arguably doing some of the same work, but using a different language, or at the least seeking a different kind of audience. Jerome Dobson, for example, has recently begun to describe the use of Geographical Information Systems (GIS) in historical geography, as a form of ‘macroscope’. This usage doesn’t come freighted with the same political claims as are current in Digital Humanities, but seem to me an entirely reasonable way of highlighting some of the global ambitions – and sensitivity to scale - that are inherent in GIS. The notion - perhaps fostered most fully by Google Earth - that you can both see the world in its entirety, as well as zoom in to the smallest detail, seems at one with a data driven ‘macroscope’. But, again, the scale most geographers want to work with is large – patterns derived from billions of data points. And again, the siren call of GIS, tends to pull humanist enquiry towards a specific form of social science. 

And finally, we might also think of the approach exemplified in the work of Ben Schmidt as another example of a ‘macroscope’ approach – particularly his ‘prochronism’ projects. These take individual words in modern cinema and television scripts that purport to represent past events – things like Downton Abbey and Mad Men - and compares them to every word published in the year they are meant to represent. 

Building on Google Books and Google Ngrams, Schmidt is effectively mixing scales of analysis at the extremes of ‘big data’, on the one hand – all words published in a single year – and small data, on the other. Of all the examples mentioned so far, it is only Schmidt who is actually using the functionality of a ‘macroscope’ effectively, making it all the more ironic that he doesn’t adopt the term. 

And almost uniquely in the Digital Humanities – a field equally remarkable for its febrile excitement, and lack of demonstrable results – Schmidt’s results have been starkly revealing. My favourite example, is his analysis of the scripts of Mad Men, which illustrates that early episodes referencing the 1950s, overuse language associated with the ‘performance’ of masculinity – words that reflect ‘behaviour’. And that later episodes, located in the 1970s, overuse words reflecting the internalised emotional experience of masculinity. For me this revealed beautifully the larger narrative arc of the programme in a way that had not been obvious prior to his work. Schmidt has little of the wider agenda to influence policy and politics evident in that of Armitage, Guldi and Weingart, but ironically, it is his work that is having some of the greatest extra-academic impact, via the anxiety it has created in the script writers of the shows he analyses. 

All of which is simply to say that playing with and implementing ideas around a ’macroscope’ is quite popular at the moment. And a direction of travel which, with caveats, I wholly support. But it also leaves me in something of a conundrum. 

Each of these initiatives, with the possible exception of Schmidt’s work, seems to locate themselves somewhere other than the Humanities I am familiar with. And this seems odd. Issues of scale are central to this. Claiming to be doing ‘big history’ sounds exciting; while claiming that more formal ‘network analysis’, will answer the questions of a humanist enquiry, appears to create a bridge between disciplines – allowing Humanists and more data driven parts of the Social Sciences to share a methodology and a conversation. But with the exception of Schmidt’s work, these endeavours seem to be privileging particular types of analysis – Social Science types of analysis – over more traditionally Humanist ones. 

In some ways, this is fine. I have discovered to my own benefit, that working with ‘Big Data’ at scale and sharing methodologies with other disciplines is both hugely productive, and hugely fun. To the extent that ‘big stories’ and new methodologies provide the justification for collaborating with researchers from a variety of disciplines – statisticians, mathematicians and computer scientists – they are wholly positive, and a simple ‘good thing’. 

And yet… I find myself feeling that in the rush to define how we use a ‘macroscope’, we are losing touch with what humanist scholars have traditionally done best. 

I end up is feeling that in the rush to new tools and ‘Big Data’ Humanist scholars are forgetting what they spent much of the second half of the twentieth century discovering – that language and art, cultural construction, human experience, and representation are hugely complex – but can be made to yield remarkable insight through close analysis. In other words, while the Humanities and ‘Big Data’ absolutely need to have a conversation; the subject of that conversation needs to change, and to encompass close reading and small data. 

The Stanford Humanities Centre defines the ‘Humanities’ as: 

'…the study of how people process and document the human experience. Since humans have been able, we have used philosophy, literature, religion, art, music, history and language to understand and record our world.'

Which makes the Humanities sound like the most un-exciting, ill-defined, unsalted, intellectual porridge ever. And yet, when I think about the scholarly works that have shaped my life, there is none of this intellectual cream of wheat. 

Instead, there are a series of brilliant analyses that build from beautifully observed detail at the smallest of scales. I look back to the British Marxist tradition in history – to Raphael Samuel and Edward Thompson – and what I see are closely described lives, built from fragments and details, made emotionally compelling by being woven into ever more precise fabrics of explanation. 

A gesture, a phrase, a word, an aching back, a distinctive tattoo. 'My dearest …. Remember when…' 

The real power of work in this tradition, lay in its ability to deploy emotive and powerful detail in the context of the largest of political and economic stories. And the political project that underpinned it, was not to ‘speak to power’, but to mobilise the powerless, and democratise identity and belonging. With Thompson’s liquid prose, a single poor, long dead framework knitter affected more change than any amount of more formal economic history. 

Or I think of the work of Pierre Bourdieau, Arlette Farge and de Certeau, and the ways in which they again use the tiny fragments of everyday life - the narratives of everyday experience - to build a compelling framework illustrating the currents and sub-structures of power. 

Or I think of Michel Foucault, who was able to turn on its head every phrase and telling line – to let us see patterns in language – discourses – that controlled our thoughts. Foucault profoundly challenged us to escape the limits of the very technologies of communication and analysis we used; and to see in every language act, every phrase and word, something of politics. 

By locating the use of a ‘macroscope’ at the larger scale, seeking the Longue durée, and the ear of policy makers, recent calls for how we choose to deploy the tools of the Digital Humanities appear to deny the most powerful politics of the Humanities. If today we have a public dialogue that gives voice to the traditionally excluded and silenced – women, and minorities of ethnicity, belief and dis/ability – it is in no small part because we now have beautiful histories of small things. In other words, it has been the close and narrow reading of human experience that has done most to give voice to people excluded from ‘power’ by class, gender and race. 

Besides simply reflecting a powerful form of analysis, when I return to those older scholarly projects I also see the yearning for a kind of ‘macroscope’. Each of these writers strive to locate the minuscule in the massive; the smallest gesture in its largest context; to encompass the peculiar and eccentric in the average and statistically significant. 

What I don’t see in modern macroscope projects is a recognition of the power of the particular; or as William Blake would have it: 

To see a World in a grain of sand, 
And a Heaven in a wild flower...
                               Auguries of Innocence (1803, 1863).

Current iterations of the idea of a macroscope, with all their flashy, shock and awe visualisations, probably score over these older technologies of knowing in their sure grasp of data at scale, but in the process they seem to lose the ability to refocus effectively. 

For all the promise of balancing large and small scales, the smaller and particular seem to have been ignored. Ever since the Apollo 17 sent back its pictures of earth as a distant blue marble, our urge towards the all-inclusive, global and universal has been irresistible. I guess my worry is that in the process we are losing the ability to use fine detail in the ways that make the work of Thompson and Bourdieau, Foucault and Samuel, so compelling. 

So, by way of wending towards some kind of inconclusive conclusion. I just want to suggest that if we are to use the tools of 'Big Data' to capture a global image, it needs to be balanced with the view from the other end of the macroscope (along with every point in between). 

In part this is just about having self-confidence as humanist scholars, and ironically serving a specific role in the process of knowing, that people in STEM are frequently not very good at. 

Several recent projects I was privileged to participate in, involved some hugely fun work with mathematicians and information scientists exploring the changing linguistic patterns found in the Old Bailey trials – all 127 million words worth. And after a couple of years of working closely with a bunch of brilliant people, what I gradually realised was that while mathematicians do a lot of ‘close reading’ – of formulae and algorithms - like most scientists, they are less interested than I am in the close reading of a single datum. In STEM cleaning data is a chore. Geneticists don’t read the human genome base by base; and our knowledge of the Higgs Boson is built on a probability only discovered after a million rolls of the dice, with no one really looking too carefully at any single one. 

In many respects ‘big data’ actually reinforces this tendency, as the assumption is that the ‘signal’ will come through, despite the noise created by outliers and weirdness. In other words, ‘Big Data’ supposedly lets you get away with dirty data.  In contrast, humanists do read the data; and do so with a sharp eye for its individual rhythms and peculiarities – its weirdness. 

In the rush towards 'Big Data' – the Longue durée, and automated network analysis; towards a vision of Humanist scholarship in which Bayesian probability is as significant as biblical allusion, the most urgent need seems to me to be to find the tools that allow us to do the job of close reading of all the small data that goes to make the bigger variety. This is not a call to return to some mythical golden age of the lone scholar in the dusty archive – going gradually blind in pursuit of the banal. This is not about ignoring the digital; but a call to remember the importance of the digital tools that allow us to think small; at the same time as we are generating tools to imagine big. 

In relation to text, you would think this is easy enough. Easy enough to, like Ben Schmidt, test each word against its chronological bed-fellows; or measure its distance from an average for its genre. When I am reading a freighted phrase from the 1770s, like ‘pursuit of happiness’, I want to know that till then, ‘happiness’ was almost exclusively used in a religious context – ‘Eternal Happiness’ - and that its use in a secular setting would have caught in a reader’s mind as odd and different - new. We should be able to mark the moment when Thomas Jefferson allowed a single word to escape from one ‘discourse’ and enter another – to read that word in all its individual complexity, while seeing it both close and far. 

I know of no work designed to define the content of a ‘discourse’, and map it back in to inherited texts. I know of no projects designed with this notion in mind. And if you want a take home a message from this post, it is a simple call for ‘radical contextualisation’. 

 To do justice to the aspirations of a macroscope, and to use it to perform the Humanities effectively – and politically – we need to be able to contextualise every single word in a representation of every word, ever. Every gesture contextualised in the collective record all gestures; and every brushstroke, in the collective knowledge of every painting. 

Where is the tool and data set that lets you see how a single stroll along a boulevard, compares to all the other weary footsteps? And compares it in turn to all the text created along that path, or connected to that foot through nerve and brain and consciousness. Where is the tool and project that contextualises our experience of each point on the map, every brush stroke, and museum object? 

This is not just about doing the same old thing – of trying to outdo Thompson as a stylist, or Foucault for sheer cultural shock. My favourite tiny fragment of meaning – the kind of thing I want to find a context for - comes out of Linguistics. It is new to me, and seems a powerful thing: Voice Onset Timing – that breathy gap between when you open your mouth to speak, and when the first sibilant emerges. This apparently changes depending on who are speaking to – a figure of authority, a friend, a lover. It is as if the gestures of everyday life can also be seen as encoded in the lightest breathe. Different VOTs mark racial and gender interactions, insider talk, and public talk.

In other words, in just a couple of milliseconds of empty space there is a new form of close reading that demands radical contextualisation (I am grateful to Norma Mendoza-Denton for introducing me to VOT). And the same kind of thing could be extended to almost anything. The mark left by a chisel is certainly, by definition, unique, but it is also freighted with information about the tool that made it, the wood and the forest from which it emerged; the stroke, the weather on the day, and the craftsman. 

One of the great ironies of the moment is that in the rush to big data – in the rush to encompass the largest scale, we are excluding 99% of the data that is there. And if we are going to build a few macroscopes, I just want to suggest that, along with the blue marble views, we keep hold of the smallest details. And if we do so, looking ever more closely at the data itself – remembering that close reading can be hugely powerful - Humanists will have something to bring to the table, something they do better than any other discipline. They can provide a world of ‘small data’ and more importantly, of meaning, to balance out the global and the universal – to provide counterpoint in the particular, to the ever more banal world of the average.

Wednesday, 11 July 2012

Place and the Politics of the Past

Preface

The talk that forms the basis for this post was written for the annual Gerald Aylmer seminar run by the Royal Historical Society and the National Archives, and was delivered on 29 February 2012.  The day was given over to a series of great projects, most of which came out of historical geography, and I was charged with providing a capstone to the event, and presenting a more general overview of the relationship between history and geography.  There was a good audience of academics, archivists and librarians, all with a strong digital cast of mind and I very much enjoyed it.  Unusually, I am also pretty sure I still agree with the majority of it even some six months after I sat down to write it.  At the same time, I have put off posting it until now because it just did not quite feel like a blog post - too much history, too much text, too much of an internal discussion among academics.   I have also recently found myself wary of blogging, having discovered (who knew?) that blogs form a sort of publication that people occasionally read.  But, the work of people like Andrew Prescott has also reminded me of just how important it is to continue having the discussion.   In the nature of a public talk, the text is rather informal, the notations slapdash, and the links non-existent.




Place and the Intellectual Politics of the Past


Currently there is a rather wonderful raproachment between historical geographers and historians; with archivists and librarians (as usual) providing the meat, gristle and spicy practical critique. This is brilliant.  These are cognate disciplines which need to be in constant dialogue.  The habits of mind and analytical tools of geographers need to inform our understanding of the past; while the mental ticks of the historian, and the authority of history as a literary genre, are necessary tools for communicating all kinds of memory to a wider audience.


Frustratingly, over the last half century, this cross-fertilisation hasn’t always flourished.  Geography Departments (and even historical geographers) have not done a lot of talking to history departments, and vice versa.  Many geographers, through the 70s and 80s, in particular, became ever more engrossed in the technical manifestations of their field; while many historians have fallen for the joys of the linguistic turn.


In part, this has been about funding and the structures of higher education.  In the great taxonomy of knowledge inherent in the creation of the notion of STEM subjects vs the Humanities, Geographers naturally gravitated towards the areas with more secure funding.  While historians, frustrated by the changing politics of their field – the collapse of ‘modernity’ as a form of historical explanation (both right looking and left looking) – pursued theory, gender and language, to the exclusion of positivist measurable change – they chose places of endless debate and disagreement – usually in the form of a various versions of identity politics - as a route to a politicised audience.  Similarly, whereas historical geographers looked to Europe and a thriving institutional base; historians tended to more frequently look westward to North America, where historical geography is almost unknown – or at least denied the security of an extensive network of independent academic departments.

 All of which is simply to say, that we are confronted with two fields that should be in constant dialogue, but which simply have not passed much more than the odd civil word in the last few decades.  In the process, they have developed different technologies of knowing and different systems of training and analysis.  It strikes me that an event celebrating Gerald Aylmer’s cross genre engagement with history and archives, with the structures of the archive, and the stories that can be told with them, is just the right place to bring these disciplines back onto the same map, or at least to reconsider where in a rapidly changing technical environment, that necessary dialogue might take place.
 
 Of course, for most of the last decade or so we have had the ‘spatial turn’ in history; and for longer than that, the creation of a post-modern geography.  Historians have struggled to define ‘space’ and context in ever more material (if still rather flabby) terms; while some historical geographers have taken theories of discourse and language seriously and extended them to the clear air enclosed by the mind-forged boundaries symbolically represented on every map.
 
 But this has been little more than a casual rapprochement – driven largely by academic fashion; and has not fundamentally changed the centre ground of either discipline. Most historians still trade in text mediated by uncertainty and theory; while historical geographers, strive to tie data to a knowable and certain fragment of the world’s surface.
 
What I want to suggest today, is that something rather more profound than the ‘spatial turn’  is also happening in the background, and that it promises to force these disciplines (and several others) back into a more direct relationship.  
 
And this change, this possibility, is being driven by technology; both in the form of the ‘infinite archive’ – the Western Text Archive second edition; and also through the direct public access to a newly usable online version of GIS-like tools, in the form of Google Maps and its many imitators. 

To deal with historians and the infinite archive first - I don’t think that historians have quite twigged it yet – though librarians and archivists certainly have - but the rise of the ‘infinite archive’ has fundamentally changed the nature of text.  It has turned text in to ‘data’, with profound implications for how we read it, and deploy it as evidence.  My guestimate is that between fifty and sixty per cent of all non-serial publications in English produced between Caxton and 1923 – between the first English press and Mickey Mouse – has been digitised to one standard or another; with a smaller percentage of serial materials thrown in.

 This has ensured that the standard of scholarship has in many ways improved.  It is now possible to consult a wider body of literature before setting out to analyse it.  But, it has also pushed us to the point where it is no longer feasible to read all the material you might want to consult in a classic immersive fashion.  Instead we are moving towards what Franco Moretti has dubbed ‘distant reading’, and towards the development of new methodologies for ‘text mining’ – or the statistical analysis of large bodies of text/data.  Stephen Ramsay’s new book, Reading Machines, illustrates four or five examples of what he describes as a new form of ‘Algorythmic Literary Criticism’, but is simply a taster for a wider series of practical methodologies.  That Tony Grafton, president of the American Historical Association could recently and hyperbolically claim that textmining was simply the ‘future’ of history, and that it was already here; reflects a truth most digital humanists (whose ranks are dominated by librarians and archivists) have been struggling with for the last three or four years; but which most historians are only now becoming aware of.

 The Google Ngram viewer, which allows you to rapidly chart the changing use of words and phrases as a percentage of the total published per year is just the most high-profile online tool in a wider technological landscape.  The associated, and ill-named, ‘culturomics’ movement being built on the back of the ngram viewer is another.  I love the ngram viewer, and spend my Sundays charting the changing use of dirty words decade by decade.  But it also forms the basis for a newly statistical approach to language.  And lest we forget, recorded language is the only evidence most historians use.

 


I am not entirely  convinced by the culturomics work, which focuses on describing social and linguistic change through consistent mathematical formulae, but in related studies by people such as Ben Schmidt, Tim Sherratt and Rob Newman, one can find the beginnings of a pioneering analysis of large scale texts that promise to remodel how we understand cultural change, and the relative influence of events.  


These graphs simply illustrate that the word ‘outside’ both grew in commonality over the course of the nineteenth century (perhaps understandably as people’s lives migrated inside); and that if this phenomenon were a reflection of naturally evolving language (embedded in people’s vocabulary in youth), its adoption according to the age of the authors whose work has been published, would look like the first graph; but that in fact it looks like the second.  In other words, these visualisations created by Schmidt suggest a history of the adoption of the use of the word ‘outside’ in response to events; and in the process give us a way of measuring the cultural impact of specific happenings – its import as measured by individual responses to them.

 
Or look at Tim Sherratt’s visualisations of the use of the terms ‘Great War’ and ‘First World War’ in 20th century Australian newspapers.  While entirely commonsensical, the detailed results of the 1940s in particular mark out the evidence for a month by month reaction to events; allowing both more directed immersive reading (drilling down to the finest detail), and a secure characterisation of large scale collections.

 
Or to bring this back to a British perspective, we can look at work Bill Turkell and I have done on the Old Bailey Online – simply charting the distribution of the 125 million words reported in 197,000 trials, to analyse both the nature of the Old Bailey Proceedings as a publication, and their relationship to words spoken in court – in this instance to illustrate, among a few other things, that serious crimes like killing were more fully reported than others in the 18th century proceedings, and that this pattern changed in the 19th century.


 

This stuff works and is important; and will necessarily form a standard component of the research of anyone who claims to understand the past through reading.  


But it points up a further issue.  If each paragraph in the infinite archive, all the trillions of words, is simply a collection of data, it immediately becomes something that can be tied to a series of other things – to any other bit of data.  A name, a date, a selection of words, or a phrase, or most importantly in this context, a place – defined as a polygon on the surface of the earth.  In other words, the texts that form the basis for western history can now be geo-referenced and tied directly to a historical/geographical understanding of spatial distribution, which can in turn be cross analysed with any other series of measures of text – textmining makes text available for embedding within a geographical frame.  


I can’t emphasise this enough: the creation of a digital edition of the western print archive means that it can be collated against all the other datasets we possess.  The technology of words, and how we engage with them, has changed; creating a new world of analysis.  With a bit of Natural Language Processing, and XML tagging; and a shed load of careful work, a component of text that hitherto has been restricted to human understanding becomes subject to precise definition: “he walked for twenty minutes from St Paul’s westward, coming first to Covent Garden, and then onwards to Trafalgar Square’, changes from a complex narrative statement reflecting an individual’s experience, into three individual locations, a journey’s route, and a rate of travel; each capable of being expressed as a polygon, a line, a formulae – a specific bit of translatable data.


What I want to say next might not sound quite right in this context.  But, I am hoping this development of text as data, and by extension, text tied to place, will have a more profound impact on our understanding of the past precisely because, for the most part, it has not emerged from historical geography.  It has been driven by people interested first in text, and only then, in data such as place. 

 As a long-time admirer of the work of historical geographers, and avid reader of it; I believe that the rise of a highly sophisticated form of desktop GIS, requiring substantial training and expertise to make work, has contributed to the evolution of a widening gap between disciplines, and has in some ways distanced historical geographers from the kind of audience historians have traditionally courted.   The rise of the geographer as ‘expert’ has been both impressive and excluding.

 But, in the last few years a real alternative has emerged.  I understand geographers are sniffy about it, and I know full well that it doesn’t provide the kind of powerful analytical environment that a fully functioning GIS Editor, Analyst and Viewer package can generate in combination with a Spatial database management system.  But it is usable and it is continually getting better.  And I mean Open Street Map, Google Earth and Google maps, and the range of open source browser side services that build on it, like BatchGeo.
Together they make available to everyone, a good and growing proportion of the tools previously only available to a technocratic elite.  In the process and in combination with the transition of text in to data; we are suddenly in a position to do something different.

 My favourite exploration of what can actually be done online in an intuitive and accessible way, is Richard Rodger’s collaborative project with the National Library of Scotland: Visualising Urban Geographies, and the associated Addressing History sites:

 

The important thing about these projects is that they allow a wider audience to use historical maps in the way a historical geographer would, and to upload their own KLM files, and to explore the data, and relate it to a modern map.  It is historical geography made user friendly.  And that is important.

 
I am also a great fan of the New York Public Library Labs, Map Rectifyer Project, which crowd sources the kind of warping of maps that people just could not previously do. 

 
And more recently, the British Library’s adaptation of the same methodology to their own map collections.

As much as anything projects like these educate a wider public (including historians) about the methodologies and issues traditionally faced by historical geographers, and generally hidden behind a beautifully presented set of final maps designed to make a point, rather than to allow an intellectual journey.  These sites form open invitations to discover all the issues associated with the underlying maps and all the problems with the data. 
 
Together, text as data and user friendly GIS make it newly possible to imagine an environment in which geographical information, and display, form a natural and unproblematic component of every other analytical process.  It makes possible a situation in which historians cease to be mere text merchants, obsessed with the perfect quote, and compelling (if largely un-evidenced) argument; and where geographers have a new access to the subtle mappings of the marks of ‘culture’ in its broadest sense – a new way of thinking about the geographical distribution of behaviours and ideas, that bring within a geographical fold questions traditionally preserved for others.  

By extension, In other words, I want to suggest that it is very much the moment for a bonfire of the disciplines, and that while history and geography can now begin to speak in new terms, the same forces are also making it possible for literature, and art history; for all the disciplines of memory and explanation, to speak in new ways to each other.  We quite suddenly share a new culture of data – and data can be translated.

I will return to both these developments in a minute.  But, by way of illustrating the kinds of things that we are now able to do as a result of this newly open and analytical framework – through the mash-up of text and space - I want to spend a little time discussing a project that Bob Shoemaker, Matthew Davies and I, and a large team of other people, recently completed, called Locating London’s Past.  Please excuse me for spending the next couple of minutes on something that sounds a little bit too much like ‘me and my database’ to be entirely appropriate.  

 

In itself, Locating London’s Past is not particularly important, but it illustrates one naïve  attempt to play with these new possibilities; to take text/data and accessible online GIS, and make something that facilitates mapping words.    

This project grew from the rich soil that is failure.  Five or six years ago, as a final component of the original Old Bailey project, we struggled to incorporate a mapping feature on to the site that could be delivered online.  But in the last few years, we realised something was changing; and inspired in particular by the Edinburgh project, we decided to try again.  The outcome - Locating London’s Past –  does three things that are new.  First, it makes available a fully rasterised and warped version of both John Rocque’s 1746 map of London; and the first ‘accurate’ OS map of the capital created between 1869 and 1880 – both of which have been fully ‘polygonised’, and related to a modern Google maps representation of London.  And second it brings together around 40 million words of text, and a raft of established datasets – a couple of hundred million lines of data - in a newly geo-coded form that can be ‘mapped’ against both area and local population, at the level of streets, parishes and wards.  And finally, it relates both these resources to the first comprehensive, parish level population estimates for the 18th century.  

In the process it brings together text and maps in a new way, delivered in a cut down Google maps container that even a historian can understand.  For the maps and GIS, we turned to Peter Rauxloh of the Museum of London Archaeological Service (MOLA), who worked with  scans and an index of place names drawn from Rocque’s map created by Patrick Mannix, to develop the kind of resource that underpins the best sort of traditional desk bound GIS project.  

The 24 sheets of the original map were turned in to a single image, and then warped onto the first reliable Ordinance Survey map from 1869-1880, creating a direct geo-referenced relationship between the first accurate modern representation of London and Rocque’s eighteenth-century version.


Rocque 1746 After Georeferencing

The geo-referencing operation involved identifying some 48 common points between Rocque's original map and a modern OS map; leaving us the task of defining all the streets, courts, parishes and wards that made up 18th century London.  In the end, this amounted to some 29,000 separate defined polygons. 

 
Parish boundaries which intersect with the street of Cheapside, London.


 
Completed street network for main area covered by Rocque's map.

Street lines expanded to polygons based on recorded width.

 
All of which gave us something rather cool – a proper, interactive and accurate map of 18th century London, that among a lot else, let’s you go from here:

 
To here:
 

And more importantly, lets you go here – all those parishes securely defined: 

 
And all those streets and cul-de-sacs:

 

In the process it makes, each parish and street, ward and cul-de-sac newly available as an analytical category – defined as a specific area, and location – defined in terms of its distance from any other place, and the route between them, its size and importance in a hierarchy of streets – and defined securely against the earth’s surface.

All of which left us with just one more task – the, to us, more familiar job of providing the text/data to put into these analytical polygons.

And for that, we brought in the material available from the Old Bailey online for the 18th century crime, from London Lives, fire insurance records, voting records for Westminster, Hearth Tax returns and plague deaths; and finally a bunch of archaeological material from Mola.   

Along with the more structured data, we ended up with a couple of hundred million words of text, primarily reflecting crime and events – descriptions of behaviours given under oath to magistrates, in court, at sessions and before a coroner; which we then processed using a combination of automated methodologies, including Natural Language Processing, and manual checking, to identify some 4.9 million place-name instances, each tied to its own polygon.

All of this data was then made available for search and mapping – including both structured and keyword searches – so both the ability to search on the crime of ‘murder’, and the word: teapot.

 

Inevitably, there are problems with the data and the map; but, it nevertheless allows us to map things like the number of small houses in the 1690s – defined as having one or two hearths as recorded in the Hearth Tax Returns.


Or more contentiously, to map the distribution of suicide cases in the coroners’ inquests found by a keyword search of 5000 inquests on ‘felo’ – as in felo de se – and ‘suicide:

 
Or the distribution of the mention of a horse, mare or gelding, in the Old Bailey.
 

We have not even begun to explore what the data tells us, nor was this created in the expectation that we would be able to do so – but the important thing is that it does allow us and everyone else, to explore this material in a new way – and do so quickly enough to facilitate the testing of new hypotheses, and random midnight thoughts.  And to quickly test words against spaces, text/data against spatial data.

Of course, all of this is contentious, and I suspect will leave historical geographers rather dissatisfied.  The original data is variable, the percentage securely geo-referenced is inconsistent and I am waiting for a few proper demographers to critique the population figures.  As worrying, the data is not currently available for the more subtle analytical approaches that have been so fruitful in historical geography.  We can’t easily define networks, for instance.   In other words, this is a rough starting point, and all the critical skills of a true sceptic are needed when using it.  But this site does allow us to play with all this data in a new way, and to come up with insights and hypotheses for further investigation.  To, for instance, map all the instances of the words for the industrial colours of  ‘blue, red and yellow’ against the natural hues of ‘brown and green’ to explore an urban environment and to suggest different ways of thinking about a wider cityscape.
 

 
 
 
It also means that we have 40 million words of geo-referenced text that we can use as the basis for a new kind of text mining – that incorporates space with linguistic change; and which will add to the geographers’ toolbox all the rather wonderful methodologies of corpus linguistics: Measures of Text Frequency and Topic Modelling to name just two.

We are, of course, nowhere near where we really want to be.  For that, we will need to have a lot more text, and a lot more subtlety.  I want to be able to map all the places in a newspaper by subject and category of article – to have a scrolling representation of places mentioned in a text as I read (either immersively or distantly).  I want to be able to use corpus linguistics, semantic search and syntactic analysis (ontologies and all the methodologies designed for text/data) in combination with both secure place name data, and historically sensitive boundary and population data.  There is not much point in comparing 18th century text with the modern road network or county boundaries; or wondering why there is not a lot of text coming from Greenland in the absence of population density figures.  I want to be able to map networks defined by individuals, defined in turn by the words they use; and networks defined by geographical measures such as road width.  What percentage of London was made up of parkland at different stages?  And what words and crimes are dominant in those different parks (Hyde Park vs Moorfields?).  And as importantly, I want to be able to test the results against secure measures of statistical significance.

All of the components to make this happen are in place – we all now work with data and data is interchangeable – subject to unending automated translation - making the main technical hurdle essentially unproblematic.  But there is still a long way to go.  And there is also a clear and present danger in the process.  And it is that danger, that I now want to turn to.

I spent last year co-directing one geographical project – Locating London’s Past - and one text mining project – Datamining with Criminal Intent.   Both projects were intellectually engaging beyond measure.  I learned more about history – and sources I already thought I knew well - doing something else with them, than I could possibly have done in a year of reading.  But I also found myself struggling against the run of the data I was helping to produce.  I came to count myself among those who Lewis Mumford had in mind in 1962 when he warned urban geographers that: 


‘… minds unduly fascinated by computers carefully confine themselves to asking only the kind of question that computers can answer and are completely negligent of the human contents  or the human results.’  Lewis Mumford, “The Sky Line "Mother Jacobs Home Remedies",” The New Yorker, December 1, 1962, p. 148

In other words, I found myself limping uncomfortably towards a positivist abstraction in which there were few people, but much data; beautiful graphs and compelling trends, but few of the moments of empathetic engagement that make history so powerful and which form a little discussed component of its authority as a genre of literature.

So, at this point, having waxed on the joys of technology and what it allows you to do, I want to stand back for a minute and remember the individual in the landscape.  And in this instance, just one individual – a man named Charles McGee or Mckay, who stood just here for over forty years, from at least 1809, until his death in 1854; making a living as a one-eyed crossing sweeper – a black Jamaican refugee from Britain’s wars of colonial expansion:

 

Or to put it differently, he stood just here, on a map created just before he arrived:

 

Or, for a map that should have included him, here:


Or if we want to get down to street level, just here – the obelisk he stood in front of, itself visible on the map:

 


MacKay became a part of the image of this cityscape almost immediately.  William Bennet was the first to record his presence, placing him before the obelisk dedicated to John Wilkes that stood at the top of New Bridge Street:

 

He was already missing one eye, but had not yet started sporting his shock of white hair:
 


A year later, in 1810, he McKay was still there:


As he was when Ackerman published the same vista in 1812.  Recognisable, even though his faced has been scratched white by some later owner of this image, clearly made uncomfortable by his presence in the landscape:

 

Five years later, in 1817, John Thomas Smith, the keeper of prints and drawings at the British Museum, gave us our first detailed portrait, and our first biography.

 

Smith claims Mckay, or McGee as he styles him, was already old beyond credibility in 1817, though another account would put his age as 50 in that year.  His hair ‘almost white’, was tied back in a tail and Smith firmly locates him at his ‘stand… at the Obelisk, at the foot of Ludgate-Hill’.  He also claims (as do most commentaries on well known street figures), that he was secretly wealthy; and that he attended Rowland Hill’s Methodist Tabernacle on Sundays; that he was lately seen wearing a ‘smart coat’ the gift of a city pastry chef, and finally that his portrait, made in October of 1815, hung in the Twelve Bells public house on Fleet Street – around the corner from the obelisk.

Two years later, George Cruickshank, includes him, broom in hand, with Billy Waters, King George III, and a host of abolitionists, in the ‘The new Union-Club’.
 



And again, in 1821, in his depiction of Tom and Jerry, ‘Masquerading it among the Cadgers’:

 
And finally, in the same year, Cruickshank includes McKay in his ‘Slap at Slop’, suggesting along the way that McKay was involved on the edges of radical London:


And so to John Dempsey’s portrait from sometime in the 1820s – which seems to me to speak of a man and a place, of a life lived in a landscape, more powerfully than any other.

 

At the beginning of the next decade, he was still there, depicted this time, from a different perspective:

 

And three years later, Mackay also became the model upon which Charles Matthews based his depiction of a modern Othello in ‘the Moor of Fleet Street’,  first performed to disastrous reviews, at the Adelphi in 1833.
 
 In the play, Mackay is depicted as engaged in a battle of jealousy and rage among the low characters of London; and is described as ‘the Moor who for many a day hath swept Waithman’s crossing over the way’ from Ludgate Hill.[i]   His spotted red bandana, clearly visible in Dempsy’s depiction, invested with gypsy lore, and gypsy power, to ‘keep woman honest, or cure the worst cold’, and given a history steeped in London’s boxing lore, and serving in the play, the role of Desdemona’s lost handkerchief. 

 The best account of his later life is from Charles Diprose’s authoritative history of St Clement Danes, where McKay lived, off Stanhope Street.  Diprose  describes McKay as ‘a short, thick-set man, with his white-grey hair carefully brushed up into a toupee, the fashion of his youth; … he was found in his shop, as he called his crossing, in all weathers, and was invariably civil. At night, after he … swept mud over his crossing… he carried round a basket of nuts and fruit to places of public entertainment...’  And according to Diprose, ‘He died in Chapel Court, St Giles, in 1854, in his eighty-seventh year.’  A later historian, William Purdie Treloar claims McKay was then replaced at his stand by a drunken soldier who ‘sometimes made 8s to 10s a day’, and drank as much each evening.

All of which is simply to say, that some people stand in the same place longer than many buildings; and have a greater right to appear on a map, than many landmarks.  As we move towards that new data rich environment of text/data and intuitive GIS; as naïve historians and the wider public, come to use the ideologically laden genre that is a map as an interface for trillions of words of text; and as they step back from their own text to view text/data from afar, I just think it is important to remember that landscapes and cityscapes only exist between the ears of their denizens – that we cannot map the subtleties of Ludgate Hill and New Bridge Street without trying to know Charles Mackay.  With Lewis Mumford, we need to ensure that we are not ‘completely negligent of the human contents or the human results’ of asking the questions only computers can answer.




[i] Note that Waithman is a wealthy linen draper with a shop on Fleet Street.  His daughter is reputed to have been especially kind to McKay, and to have received a legacy from his on his death of £7000.  Treloar gives a more detailed account of Waithman’s role as alderman and MP, and suggests his daughter regularly took out soup and warm food to McKay, p.124.
.