When the very first digital images from medieval manuscripts were posted on the web way back in the age of dinosaurs (late 1990s I think it was) the potential for such resources for scholars and hobbyists was apparent. Also very apparent was the massive scale of the enterprise in terms of labour and expense before anything like a useable academic resource was available. Web technology was less advanced, so images tended to be either rapid loading and crude, or very slow loading and high quality. Then there was the issue of what those who put the images up thought the users ought to be able to do with them. In many cases not much. My favourite (not) message was "You may not download this image." Yeah, OK, so how come I'm looking at it?
When I first put up Medieval Writing around 2000, I attempted to include links to all sites with reasonable displays of digitised manuscripts. It was actually possible. Since then it has become harder and harder to keep up, and in the last few years there has been an avalanche of sites with complete digital facsimiles of such a quality as would enable various kinds of fairly detailed research.
There seemed to be certain national characteristics in relation to the display of manuscripts on the web. German sites went for complete facsimiles with grungy interfaces. French sites went for beautiful displays of miniatures in thematic arrangements, but isolated from their texts. French provincial libraries displayed their treasures, often buried in municipal government websites, with complex and changing urls. Switzerland went in for the full bells and whistles, complete manuscripts, funky interface and open access with its splendid e-codices website. I believe a Scandinavian site from the Helsinki University Library was the first to claim its medieval manuscript images could be used copyright free, but that resource seems to have vanished. British institutions, on the whole went for selections of pages and illustrations, so that it was never possible to study a complete manuscript. And if you could, they told you not to. American libraries flaunted their translocated treasures with arty selections.
Now the race is on to make the world's manuscript treasures available. And institutions are realising that they can allow people to make use of them, because for starters, how are you going to stop them, and for seconds, what possible harm can it do to the original? The circulation of CDs of dubious provenance sold on eBay from countries like Spain of material on the web claimed to be copyright has been followed by a crowd of scholars republishing and re-referencing material on blogs, tweets and other new media. Even the British Library now allows its digitised manuscript images to be used without copyright restriction for non-profit purposes. Now if they would only digitise a few more complete manuscripts ....
While talk of persistent urls for specific manuscript pages has been going on in library circles for a long time, some manuscript sites are finally making use of them, allowing other users to construct databases on whatever themes they choose across various sites. But it is still all taking time. Every listing of digitised manuscripts on the web is incomplete or out of date.
Now one Giulio Menna of the splendidly named Sexy Codicology website is plunging into the task of providing an interactive map of all libraries in Europe offering digitised manuscripts on the web. Check the progress here. Best of luck Giulio. Big job. It's an independent project and a labour of love. The guy is crazier than me.
A companion to the website Medieval Writing, concerning itself with medieval handwriting and its cultural setting, now expanded to encompass aspects of medieval heritage and material culture. Tweeting as Hipster Bookfairy . Gradually putting medieval photos on Flickr
Showing posts with label copyright. Show all posts
Showing posts with label copyright. Show all posts
Friday, July 12, 2013
Sunday, May 29, 2011
Reprise on Google, eBooks, Copyright and All That Jazz
An enigmatic personage by the name of Dr Beachcomber has sent me an email with a link to his posting Google Burns the Library at Alexandria. He has included my reply as a comment on his blog, so I am returning the compliment by referencing him here. Is this what you call some kind of hippy blog-in?
While I have mildly chastised him for over dramatics in headline writing, the books not actually being burned as a result of having been digitised, there is a issue of concern regarding the quality of scanned digital editions, and another issue brought up by another commenter on the recopyrighting of material already in the public domain as a result of it being reprinted or republished. There is also the very tricky issue of the destruction of original printed or written material after it is digitised.
Taking the last first (Hey, I'm in Australia, we are upside down here!), I was many years ago doing a research project which involved examining museum records and objects. Now museum curators have a habit of updating their records when they think that a person looking at them is some kind of expert and they ask them questions about things. For historical reasons, I wanted to know what the original records said about the objects. With the old handwritten cards and registers, it was possible to separate the original records from later annotations, and even to work out who had made the annotations and when. Only one museum had an electronic catalogue at that time (1991), and they were quite disappointed that I actually wanted to look at their tatty old paper records. I guess the question is, how many old paper backups do we need to keep for safety? The same applies to books.
On the second issue, I was told many years ago by a copyright legal bod in my university that it was legal for me to scan out of copyright visual material and republish it digitally, but it was illegal to reproduce digital scans from modern facsimile editions of out of copyright material. My only question about that is, how could anybody tell? At the moment the business interests are noisily defending ever increasing copyright restrictions, but the ready availability of copying and reproduction technology is going to make soup of that, and real soon. I suggest that if you have some favourite old, genuinely out of copyright, books in your particular area of interest or expertise, digitally reproduce them yourself, circulate them among your friends and colleagues, and loudly announce them as public domain.
The quality of some of the old material scanned and placed in the public domain is an issue. Dr Beachcomber is determined that Internet Archive editions are better quality than those from Google, but I bet he has never spent three days printing a long book page by page from two separate Internet Archive scans, hoping that the pages missing from the two editions do not actually coincide at any point. The end result was a largely black and white edition with occasional colour pages, none of which had bookmarkable or cut and pastable text as they were simply image scans of pages. And the Kindle editions are similarly unnavigable and messily formatted. And the text only versions are unformatted to illegibility and full of OCR errors. But apart from that they're alright. I suspect that there is just some degree of luck with the digitisation of particular works, and how carefully they have been done.
I have touched on these issues in earlier posts, Eeee! Books, and Scribes, Copyright, Crime and Google, with a short note at the end of Horrible Old Handwriting. I guess the whole issue is just not going to go away real soon.
The whole issue of preservations of books and text is, of course, not new, but there are so many texts to preserve these days. We have almost no original Roman era texts of the Latin Classics, because they were written on papyrus rolls which fell to bits. These works are mainly preserved from much later copies in vellum codices, much more durable, produced by Christian monks. The thought of these celibate ascetics solemnly copying down the erotic poetry of Ovid and the like is always good for a giggle, but they did. There have even been conspiracy theories that the monks actually forged all the Latin Classics. I doubt it, but how much did they edit, correct, annotate and standardise these texts? Perhaps Cicero or Livy might be surprised to discover what we think they had written.
Postscript: With apologies to Dr Beachcomber, after rechecking, it seems that the download I had such trouble with was a Google scan, although I accessed it through the Internet Archive. It was one of a large set uploaded by one tpb, who seems to be a very messy worker. Perhaps I was dead unlucky, because there appears to be another edition of the same book available through the Internet Archive which is not from Google, so at least they are not claiming a monopoly for their grotty scans.
While I have mildly chastised him for over dramatics in headline writing, the books not actually being burned as a result of having been digitised, there is a issue of concern regarding the quality of scanned digital editions, and another issue brought up by another commenter on the recopyrighting of material already in the public domain as a result of it being reprinted or republished. There is also the very tricky issue of the destruction of original printed or written material after it is digitised.
Taking the last first (Hey, I'm in Australia, we are upside down here!), I was many years ago doing a research project which involved examining museum records and objects. Now museum curators have a habit of updating their records when they think that a person looking at them is some kind of expert and they ask them questions about things. For historical reasons, I wanted to know what the original records said about the objects. With the old handwritten cards and registers, it was possible to separate the original records from later annotations, and even to work out who had made the annotations and when. Only one museum had an electronic catalogue at that time (1991), and they were quite disappointed that I actually wanted to look at their tatty old paper records. I guess the question is, how many old paper backups do we need to keep for safety? The same applies to books.
On the second issue, I was told many years ago by a copyright legal bod in my university that it was legal for me to scan out of copyright visual material and republish it digitally, but it was illegal to reproduce digital scans from modern facsimile editions of out of copyright material. My only question about that is, how could anybody tell? At the moment the business interests are noisily defending ever increasing copyright restrictions, but the ready availability of copying and reproduction technology is going to make soup of that, and real soon. I suggest that if you have some favourite old, genuinely out of copyright, books in your particular area of interest or expertise, digitally reproduce them yourself, circulate them among your friends and colleagues, and loudly announce them as public domain.
The quality of some of the old material scanned and placed in the public domain is an issue. Dr Beachcomber is determined that Internet Archive editions are better quality than those from Google, but I bet he has never spent three days printing a long book page by page from two separate Internet Archive scans, hoping that the pages missing from the two editions do not actually coincide at any point. The end result was a largely black and white edition with occasional colour pages, none of which had bookmarkable or cut and pastable text as they were simply image scans of pages. And the Kindle editions are similarly unnavigable and messily formatted. And the text only versions are unformatted to illegibility and full of OCR errors. But apart from that they're alright. I suspect that there is just some degree of luck with the digitisation of particular works, and how carefully they have been done.
I have touched on these issues in earlier posts, Eeee! Books, and Scribes, Copyright, Crime and Google, with a short note at the end of Horrible Old Handwriting. I guess the whole issue is just not going to go away real soon.
The whole issue of preservations of books and text is, of course, not new, but there are so many texts to preserve these days. We have almost no original Roman era texts of the Latin Classics, because they were written on papyrus rolls which fell to bits. These works are mainly preserved from much later copies in vellum codices, much more durable, produced by Christian monks. The thought of these celibate ascetics solemnly copying down the erotic poetry of Ovid and the like is always good for a giggle, but they did. There have even been conspiracy theories that the monks actually forged all the Latin Classics. I doubt it, but how much did they edit, correct, annotate and standardise these texts? Perhaps Cicero or Livy might be surprised to discover what we think they had written.
Postscript: With apologies to Dr Beachcomber, after rechecking, it seems that the download I had such trouble with was a Google scan, although I accessed it through the Internet Archive. It was one of a large set uploaded by one tpb, who seems to be a very messy worker. Perhaps I was dead unlucky, because there appears to be another edition of the same book available through the Internet Archive which is not from Google, so at least they are not claiming a monopoly for their grotty scans.
Thursday, June 03, 2010
Scribes, Copyright, Crime and Google
Now I know I do like to rabbit on sometimes about the continuities and discontinuities in written communication in the middle ages and today, but during the course of the last day or so I have truly found myself, like Alice, down the rabbit hole and behind the looking glass. It all started when I found a link to an interesting old French paleography book published in 1892.
In the days of medieval scribes, there was no such thing as copyright. No sooner had an author put away his quill than everybody was free to transcribe his words, paraphrase them and incorporate them into new contexts. It was only the industrial production of books which set up the conditions for protection of authors and publishers, and then not for some centuries. The purpose of copyright was not to inhibit the dissemination of words, but to encourage them by protecting the investment of those who had set up print runs of books. Authors got paid royalties, so they didn't have to rely on the patronage of kings and magnates in order to eat.
Many interesting books published long ago are no longer available, often because they are only interesting to a small number of people, but interesting nonetheless. Google has been collaborating with some very eminent libraries to make these available again through Google books, but there is a catch. Copyright laws are not the same the whole world over. Rather than try to untangle the mess, Google has simply made certain books unavailable in full text to countries outside America if they have been published between around 1870 to the 1920s, whatever their actual copyright status. This was the case with this old paleography book, written in France, which I was trying to access from Australia.
Trolling around the web to resolve this issue, I discover that one suggestion is to obtain a free proxy in America, so that Google thinks that is where you are. This is very easy. There are hundreds of them, and they make up new ones every day as the old ones are shut down. Furthermore, they advertise this service in terms of not allowing your web surfing to be tracked, and enabling you to access sites banned in your school, workplace or country of residence. The sites have a tendency to have "up yours" or "in your face" kinds of names. I have also discovered that this is the very easy fix to our very stupid country's very stupid proposed mandatory internet filtering. So here I am, consorting with pornographers, gunrunners, terrorists, clandestine Facebook users in the workplace and God knows who else, just in order to read a very old academic book which actually is thoroughly out of copyright, here and everywhere else. Furthermore, very respectable archivists and academics have shown me how to do it, because it is not illegal to read or download this book. It's just the kind of company you have to keep in order to do it.
Of course, there was a catch. The proxy would not allow me to download the book as a pdf file because it was too big. Another proxy with a large download allowance was not actually in the USA. So the second suggestion was to see whether it was on the Internet Archive as a text. It was, but ..... if you clicked on the link to download the pdf, you got sent to .... Google books. I am now part way through the process of printing a large book one page at a time from the Internet Archive, because it is one of those reference type books that are useful to have to hand. It's got huge bibliographies and a large dictionary of Latin abbreviations. Yawn! But it is still cheaper than flying over to Harvard to look at it in the library from which it was Googled.
Isn't it about time that the publishing industry got over its paranoia, and there was some means of releasing elderly books of specialised interest without hysteria about copyright? If the book is out of print, it should be available. The big publishers are not going to have their sales of the next J.K. Rowling or Dan Brown knocked off by a few harmless oddballs downloading dictionaries of medieval Latin abbreviations. I'm sure that long deceased authors of specialist academic material would be fascinated if they could know that somebody still did want to read what they had written, just as I'm sure that even longer deceased medieval scribes would be fascinated and bemused by people rescuing scraps of their work from the bindings of later books and poring over them.
Meanwhile, if you don't hear from me again, you will know what has happened. "Knock, knock!! "Ello Sunshine, you're nicked! You've been using HideMyAss in order to download a little cursiva bastarda. Just come along with me, Madam."
Tuesday, January 19, 2010
Copyright and Old Stuff
Do you ever get the feeling that the whole issue of copyright is completely out of control? The ease with which things can be reproduced, and the various media in which they can be reproduced, have led to endless churning debates in which nobody really seems to have clear, legally sound answers. On the one hand there are the anything goes brigade, who seem to think that because an object is old, any reproduction of it should be copyright free. On the other hand, there are publishers and custodians of material who seem to believe that they have rights over any reproduction of anything that they have ever owned, or anything that resembles anything they have ever owned.
There are so many hypotheticals that can reduce the debate to a shambles. For example, if I decide to put a picture of my living room on my social networking page, and I happen to have a painting by a living artist or a published print on my wall, am I supposed to pay them a royalty? If I take a picture of a major heritage monument from the same place and in the same weather conditions as that in a coffee table book, will they accuse me of piracy? On the other hand, if I take their picture and work some digital jiggerypokery on it, will they hunt me down for pinching the source material, and anyway, how would they know, given that it is a large, public, inert object?
The issue arose with me recently concerning some illustrations of museum material, which had nothing to do with medieval manuscripts as it happens, in which a publisher asserted that illustrations of museum objects were copyright to the museum and permission had to be sought to publish them. Now as it happened, those illustrations were drawings derived from photographs which I had taken myself, but as I had taken the photographs in the museum under the condition that I sought permission if I were ever to publish them, I had actually sought that permission. However, to my way of thinking, that is not copyright, that is contractual obligation, not to mention common politeness. I believe that is an important difference, as I would seriously doubt that ancient objects themselves can be copyright.
I have had some occasional correspondences with libraries, and with users of Medieval Writing, over this issue in relation to the reproduction of medieval manuscripts. There are some who believe that because manuscripts are old, that they are not copyright. However, the photographs of those manuscripts may be subject to copyright restriction, and libraries may place conditions of use on photographs which are either purchased from them or taken with their permission within their walls. For photographs published in books, that is covered by copyright. For photographs taken by a user or purchased from the institution, I would assume that, like the museum objects, that would actually be covered by contractual obligation.
However, photographs have been around for some time now, and I remain quite unclear about the copyright status of old photographs found in somebody's bottom drawer, which they have handed on to me because they thought I might find them useful. I remain unclear about copyright claims that are couched entirely in the terms of print media when internet reproduction is different in so many ways. And I remain unclear about the actual rights of museums and libraries over the objects and their representation, as opposed to the reproduction of those objects under conditions which are clearly specified by copyright or contract.
I try to work within what is legal, and fair to both curators and users. There are a number of very important libraries and archives now that are putting up very impressive digital editions of manuscript material on the web, free for all to use. This makes material available to scholars and interested parties who might not otherwise be able to get access, and it does aid conservation by reducing handling of the originals, but it does cost money. What needs to be avoided is putting this material into the hands of corporations which have the objective of making profits, not increasing access to cultural heritage.
Meanwhile, the latest edition to the website is a script sample and paleography exercise of a bit of 13th century Gothic textura, full of speculative historical romance and devoid of copyright issues.
Subscribe to:
Posts (Atom)