Sunday, April 13, 2014

Thoughts on Provenance




            Well, no report on the conference I was to attend.  Didn’t get to go.  Got sick thanks to germs passed around at a family party. I’m still bummed because I was really looking forward to going.  There’s another conference in New Orleans that I may attend.  Will keep you posted.

            Since I could do nothing these last 10 days, but cough, I’ve spent the time thinking (that is when I wasn’t watching a marathon of Indiana Jones movies).  It occurred to me that although the importance of maintaining provenance and collection integrity is obvious to me it doesn’t appear to be obvious to many repositories.  Granted most of the dismantling of collections was done in the past, at least at the libraries where I work, but it doesn’t seem even today that many library programs emphasize an understanding of provenance as part of their curriculum.  So much information is lost when the history of a collection is lost due to dismantling.  Perhaps an example will help clarify.  As I’ve talked about before I have been working on a photograph collection of one of the universities.  I noted that the photographs had all been literally dumped together.  Not only was much of the original order lost, but also any history of the photographs or the photographer.  What is known is that the photographs were taken to record the events and people of the university.  What is not known now is which department initiated which group of photographs.  Some were for the yearbook.  (Obviously some of the portraits were taken for that purpose.)  Some others were for the campus newspaper.  The public relations department maintained photographic archives and used the photographs for various publications. Portraits, especially of board of trustee members, were part of their collection it seems.  The alumni association also had photographs of various events and people.  Some of the photographs were probably from different campus organizations and social clubs apparently given by individuals or the club itself. In other words where the photographs originated in most instances is based on conjecture.  Most of that information is lost.

            All we really know is that a particular photograph was generated by some department or individual connected to the university.  Just think how much more meaningful it would be to know that a set of photographs were taken as part of a fund raising project or for a newspaper article about an important event or person at the university.  The metadata (information records) about that photograph would be so much richer because more of the context of the photograph would be known.  Why a photograph was taken and for whom can be as important as what it depicts because it provides much of the history not just of a particular photo but also of the institution and how it functioned at a certain period of time.  In other words, maintaining provenance is important even for a bunch of photographs.  

Wednesday, March 26, 2014

Digitizing photographs - End of Phase 1


           Well, we’ve almost finished the first 2000 photographs to be sent off to the University of North Texas for digitization.  It wasn’t so bad with three of us on the project.  What have I learned? First off you need to edit and re-edit your work.  Metadata and inventories are monotonous so mistakes are almost inevitable at least mine seem to be.  If you’re working in Excel you have to be especially careful because, of course, there is no spell check. That’s another reason to dislike Excel. Another is some of the automatic features. They drive me crazy. It’s fine for spreadsheets, but tough for inventories.  Just because the first two words are the same doesn’t mean all the rest an entry is.  I’ve done all of the inventories before this in Word. It has its problems too. Numbering and indenting are two of the most annoying.  All of a sudden the numbers change, indents won’t align. What’s up with that?  I try to remember to disable the automatic features but I don’t always.  Trying to do an outline is often a nightmare.  Even with that I prefer Word to Excel for inventories.

          What other things might I consider doing differently? Well it would be nice to have an expert on the collection around to help us identify individuals.  We did have a book of the history of the school and that was great and, of course, the internet is wonderful if you have enough of a name.  One area where we needed more assistance was with dates.  Most photographs were not dated, many not identified.  We had to guess dates based on clothing or the absence of computers and that sort of thing or leave the date cell blank.

          The other thing that I might consider next time is to expand the metadata.  We had no column for the type of photograph – black and white, color, slide, or whatever.  Another category could have been photograph size – 8x10, 5x7 or in metric.  That’s often helpful for a researcher.  UNT may do that. I am not that familiar with their metadata schema and how much if anything they add to what is sent to them.  I guess I need to find that out. 

           The last problem I see is an ongoing one. I have mentioned this in other posts.  How do you choose the photographs to digitize objectively? I still don’t have an answer for that one.  Portraits are the most difficult.  They are already on line in the yearbooks so if you have a date and name you can find someone’s photograph.  Often the scanned yearbook or newspaper image is not that high a quality so that may be one reason to digitize the portrait separately.  I guess what concerns me most is that not every faculty and staff at the university had photographs in the archives.  Actually it was pretty haphazard.  For example, the seventies and the nineties were fairly well represented, but not the eighties or the earlier years.  The university or any creator of a collection needs to impose more order on their collections to avoid this problem.  Like I said before identify your photographs, date them, and put them in some organizational scheme. You will make an archivist very happy.

        Off to a conference next week.

Sunday, March 23, 2014

Digitizing Photographs - Week 3


            Well we are slowly making progress with item level description (metadata) for our photograph collection. Yea!!! Our new approach of a mini-assembly line seems to be working. I’m still choosing photographs, re-housing as needed, and I started doing a little research on the individuals in the photographs. That’s making the metadata go quicker and allows my partners time to do a more in-depth research on some of the photographs.  That’s paid off because they have found name errors and were able to indicate the correct name (Jorge, not George, for example). 

            The problem as I see it is the subjective nature of picking photographs that should be digitized.  As I talked about before we have criteria for choosing what to digitize, but whether that will meet the needs of researchers we have no way of knowing.  In some respects we are doing an on line exhibit that gives a sample of what the collection contains.  At this point we have no way of determining the possible uses this collection might have – genealogists, alumni looking for friends, history researchers.  Perhaps in the future we can devise some test of what photographs are used and why – the way a museum tests to see the value of a exhibit to its patrons. 

            I do think one important consideration is the precision and accuracy of the search engine.  At first we did not choose a photograph if it had been in the campus newspaper or yearbook. The exception were the older photographs from the twenties.  We are rethinking that.  Although the photographs may be up on line unless you know exactly what publication they are in and where in that publication, the search engine is not able to find them.  For example, college catalogs have photographs of activities around campus.  Rarely do these have the individuals identified.  When we digitize the same photographs we can add names and dates and location as part of the metadata so the search engine can see it.  An example is a group photo of a singing group. In the university catalog the group has no identifying information.  Our metadata does.  It’s something to consider if you have similar situation.  If it’s up on line, but a search engine can’t find it that doesn’t help anyone.

Sunday, March 9, 2014

Getting Ready for Digitization


           Last blog post I talked about the initial processing of a large photograph collection to the file level.  The goal as I noted was to provide the university with some intellectual control over the collections in their archive. That really is the point of processing, that and helping to preserve the physical material. After the project concluded several collections were earmarked for digitization by the library if and when money became available. Well, money has become available so we are starting to prepare the photograph collection for digitization.

             As I mentioned there are at least 10,000 photographs processed to the file level, not the item level.  That doesn’t count negatives or slides.  I emphasize that because digitization requires metadata (information) about each photograph digitized.  That means we must process every file to the item level and we have only a few weeks to accomplish this. Now I had spent the better part of a summer organizing the photographs to the file level and imposing some order. My goal was to make that photographs accessible to anyone at the university who might need photographs from a particular topic like university buildings or football or faculty.  I wasn't processing for digitization.  As I mentioned before, the photographs had arrived at the library thrown into boxes, some in manila folders, some in the envelopes from the printing company, but most just simply tossed together.  I should note that a previous attempt had been made to identify the individuals in the photographs, but this had failed and those photographs were simply thrown into boxes for another move.  Since the photographs were for the most part kept by the Public Relations Department many had been used in publicity.  Others had been published in the yearbook.  The yearbooks and university newspaper are already digitized so many of the photographs are already on line.  Of course unless you search for every photograph there is really no way to know what is already on line.  Money is limited so that wouldn’t work.  What to do? Well, you compromise and do the best you can with what you have at least that is what we are doing.


Some of the more organized boxes prior to file level processing

              Since there’s not money to digitize everything we had to make decisions.  Older photographs where the image could be identified were chosen because of their importance to the early history of the university. Even if the photograph was online in yearbooks or the campus newspaper, these early photographs will still be digitized.  The rationale is that a digital image from the original photograph would be clearer than one in a yearbook picture.  Attempts are being made not to digitize duplicate pictures, but this has proved difficult because the same picture may be in multiple files.  Portraits of significant university presidents, for example, can be found in various files.  One person working on a collection might catch duplication, but with multiple people helping it is impossible. I’m not sure how to avoid duplication given time and money constraints.  This is one area where we are still addressing, especially in terms of portraits.  Even if you avoid duplicates how many different portraits of one particular person do you need? If it was a faculty member there might be a portrait for every year they taught and that may have been years.  

             At first we started with everyone taking a box.  Each individual made the decision, which photographs were to be digitized, numbered them, and provided the metadata.  Progress was slow.  Currently we are approaching the problem like an assembly line.  One person, me, goes through each box, chooses the photographs to be digitized, numbers each item, and re-houses as necessary (most of the photographs are not in their own sleeves as they should be).  The next person is in charge of entering the metadata and making the final decision of what gets digitized.  Hopefully this will better address the duplication issue and allow the proper housing of the material.  We’ll see how fast it goes.  We have also decided to divide the collection in two, that is, not try to do it all at once.  We only have money for 1500 photographs and last count we were near a thousand.  Wish us luck.

Wednesday, March 5, 2014

Dealing with the Real World: Archival Processing Compromises



                        I’ve talked about original order and provenance before and noted that  problems result when these rules are ignored. Once collections are divided provenance can be lost.  Once original order is ignored, the organizational scheme of the creator is lost. Basically trying to recreate what was changed often takes too long and may cause more problems.  You must compromise.  For archivists the goal is to expedite processing to enable accessibility and gain intellectual control over the collection.  Even if you could undo well-intentioned destruction of original order budgetary constraints and time often provide limitations.  The collection that I am working on is a case in point.  It is photographs.  The collection arrived at the library from the Public Relations Office and the Alumni Association although it is not really clear who sent what. The original creator department is not known as the records have been passed around and stored hither and yon throughout the university.  So we don’t know who collected the photographs and we don’t know the organizational structure that was used because that’s been lost over time. Some of the photographs are numbered although exactly what the numbering means is not altogether clear. Some of the photographs arrived at the library when a building on campus was being renovated.  The rest came from a retiring staff member’s office.  Unfortunately she was a saver, but not particularly well organized.  Boxes from her office had little to no organization with photographs and unrelated papers mixed together.

An example of one of the smaller boxes of miscellaneous material

                        I spent the good part of a year trying to impose some order to the records and attempt to find any underlying order that might still exist.  First step was to accept that that the original order could not always either be determined or be restored.  The next was for the library to make some decision about what they would preserve and what they could not.  As we talked about before, not everything should and can saved.  Our problems were acaerbated when the archives had to be moved again because the room where they were housed was needed as a classroom. Time was short.  Did I mention that money was also limited particularly for archival storage material?  The best we could do were a few archival sleeves, but we did have archival folders and acid-free boxes.  It was a start towards preservation and organization, but definitely a compromise.

                        The first step in dealing with a collection like this is to do an appraisal and come up with a processing plan.  That required looking into each box and trying not to become too overwhelmed. No one had gone through the photographs to weed out those that were not particularly good. Most had no identification.  Some had been used in previous publications or appear to have been.  Some were professional photographs.  Some were simply candids  - some good and some bad. Did I mention that there are over 10,000 photographs not counting negatives and slides?  As part of the plan that was developed the library staff made some decisions of what to keep and what to discard.  For example, yearbook photographs were already digitized so they were not kept nor were poor photographs (out of focus, head chopped off, poor lighting, lots of pictures of unidentified Homecoming bonfires – that sort of thing). What organization that could be determined was kept.  For example, there was an entire box labeled “Social Clubs.” This became one of the sub-series under a series called “Students.”   Examples of the series developed include the following: Buildings; Athletics; and People.  People had sub-series of Faculty, Board of Trustees, and so forth.  Each series has files, such as early buildings or current buildings under the series “Buildings”; for the series "Athletics" the files are football or baseball, etc.).  These series seemed to have been part of the underlying order as best we could determine.

At least the outside box was labeled even if the photographs weren't
               The best way to approach a mess like this is one box at a time.  At least that’s what seemed to work for us.  I think that is the approach to use when sorting through papers at home too.  It’s easy to get overwhelmed.  We had one of an  emeritus professor help us with some identifications.  He wasn’t emeritus enough to be able to identify everyone, but it helped. The next step is digitization and that means item level numbering and metadata (information) and I hope best practice collection care. Oh help! Must remember one box at a time.

Saturday, February 22, 2014

Oral Histories


            Oral histories are among my most favorite type of collection.  I’m fortunate because there are tapes at each of the institutions where I work so I’ve gotten to do a lot.  Some are really fascinating while others are just ok. A lot depends on the interviewer.  When I do transcripts, I make a verbatim transcript so as you read it you have a feel for how the person speaks.  Verbatim transcripts can be tedious to do because often the quality of the tape is poor, or people mumble or talk over each other.  It can be challenging.  I should have studied court stenography.  That would help, except then I think you have to transcribe the shorthand notes – twice the work.  Guess my way is best.

          The first step in doing an oral history is to make a copy of the original.  You’ll use the copy to do the transcript so you don’t risk damaging the original.  If you have an old audiotape there are still machines around that can make copies of tapes.  Of course, then you have to have an old tape recorder in order to listen to it.  The other alternative is to digitize them to CDs.  I have done both.  Once you have a copy you’re ready to go.  One of the librarians I work with suggested an application called the Amazing Slow Downer (http://amazing-slow-downer.en.softonic.com/).  It’s also available on ITunes.  This app was designed for music, but works great for CDs of oral histories.  You can adjust the speed of the speech as well as the bass and treble tones.  All of that can help with the transcription.  You can try this app for free. To buy it, I believe, is around fifty dollars.  I did fine with just the trial.

            Have fun with oral histories. There’s no telling what you will learn about the past. 

Monday, February 10, 2014

Preserving paper - more information


          Photographs can be preserved by keeping them in cool, dry place out of light. Don’t forget to label them appropriately.  Now what about paper? How do you keep your important papers whether they are personal papers or those in an archive? Well, what matters most is the type of paper that you have.  If you are a big newspaper clipper, expect anything you save made of newsprint will turn yellow and brittle in a pretty short amount of time. Why?  Well newsprint is an example of cheaply made paper made from wood pulp. Wood has lignin, which is an acid.  It is this acid in newsprint that causes the deterioration.  On the other end of the spectrum is bond paper.  Bond paper is made from cotton rags, which do not have acid.  They will last 100 years if kept out of the sun and in a cool and dry place next to other non-acidic papers.  That's why colleges have historically demanded that masters theses and dissertations be printed on bond paper because they want the material to last as long as possible.  Other materials like archival non-acidic file folders or other non-acidic products have had the acid (lignin usually removed) to extend their shelf life.  They can help reduce the migration of acid from acidic paper to non-acidic material.  Did I mention that acid will move from one piece of paper and chemically alter material next to it? Test this by putting a news clipping on a piece of paper and leave it in direct sunlight for several weeks.  When you lift the news clipping a brownish stain (acid) will be left on the paper underneath.

            So what can you do to preserve your material?  Not much if it is newsprint.  Archives microfilm their newspapers.  Microfilm is stable and will last for about 100 years.  The other recourse is to copy the article onto bond paper and throw the clipping away.  If you must save the clipping use acid free paper or other archival material between the clipping and other papers.  The clipping will still deteriorate but you will protect the surrounding papers at least a little. The other way to protect paper is to reduce human handling.  Some archives require white clean gloves be used when handling paper. I find I do more harm than good with gloves. Do wash your hands. No drinks or food around paper material you want to preserve.  No rubber bands to bundle paper together.  Rubber bands deteriorate, turn black and may stick permanently to the paper.  Metal paper clips and staples will rust over time if there is moisture in the air.  Use plastic paper clips.  You can keep paper together by folding a blank piece around the group you were going to staple.  Or go ahead and staple if you must just know that it may rust over time.  Other considerations – Store paper in archival boxes  or in metal filing cabinets.  Never store paper products in wooden containers – acidic, remember? Archives use special non-acidic shelving for their collections.  Coating on that shelving does matter.  Baked coating is recommended.  For archives that have wooden shelving and can’t afford to replace it, a cheap reasonably effective solution is to line the shelves with archival board (a non-acidic cardboard only available from archival supply houses).

            The newest solution to preservation is, of course, digitization.  It works, but we’re just not sure for how long.  I have files in Excel, for instance, that can only be read using old Excel software.  The newer versions don’t support the older ones. You must keep upgrading your files to stay current and I wasn't quick enough. (Of course, there are "how to upgrade" these files on line. I should do that but haven't.)  For now saving as pdfs or in Word are the best options for long-term usefulness.  For photographs JPEG and TIFF have stayed readable so far. 

For more information on paper preservation see:

Information on archival shelving
http://www.archives.gov/foia/directives/nara1571.pdf   “Use a powder-coating system to paint all painted metal shelving surfaces (including map cases, museum cabinets, etc.) used within all records areas. The powder-coating polymer must be a polyester epoxy hybrid or best equivalent available that passes NWT- conducted or independent lab tests for hardness, coating stability, bending, coating adhesion, and coating durability. The paint must not exceed the off-gassing limits specified in Appendix B. Do not apply powder coating to the metal surfaces onsite in the storage area.”