Showing posts with label Digital indexing. Show all posts
Showing posts with label Digital indexing. Show all posts

Wednesday, April 24, 2013

Is "Digital" Dead?

It occurred to me while I was working on the project "Oral History in the Digital Age", that I wonder what the inevitable post-digital age will be characterized by... I'm wondering if at some point we'll be able to stop putting the word "digital" in front of everything and just call it "the work we're doing now". This stems from a question I've had for several months now: what does "digital" really mean? Despite it's flagrant use as modifier for all sorts of things, I think it just means we use computers, which is such a given at this point in history that we hardly need to make a note of it. I'm starting to think it might be time to reign in that term, and make sure that our enthusiasm for the potential that computers bring to us (and the potential of web-based social connectivity) does not get confused with what our actual work is.



Wednesday, April 3, 2013

Timecode Metadata



Timecode metadata are the critical link in the between textual content and audio or video in digital environments. Different architectures for timecode deployment have evolved independently in the creation of digital oral history collections, and all help to significantly increase digital accessibility. With many models now on the table it is an appropriate time to take inventory of what approaches are available, closely evaluate the relationship between these models, understand the range or textual data they are linked to, and elucidate the current “state of the art” to find common ground for future developments. 

Timecodes are being put to use in two broad ways: 1.) as transcription timecodes, enhancing full text transcriptions with a cross-reference to time points in the source audio or video, and 2.) as audio or video file metadata enhancing a longer audio or video file, or A/V timecodes. Within A/V timecodes two basic models are emerging, one that uses timecodes pointing to a single point in time in the digital file, allowing the user to play forward from that point. (We might call these indexing point timecodes.)  In another model, (which we might call passage timecodes), timecodes are defined as inpoints and outpoints giving meaningful content within a longer digital file its own begining, middle and ending. 

The latter model of defining passage timecodes can take place in database environments where the in/out points are just references that move the listener digitally (hypertextually) to the passage of interest. In other contexts, practitioners manage oral histories by hard-editing passages permanently, thus creating segments or clips from the full length digital source file.   

All timecode deployments require choices to be made--regarding the frequency of transcription or indexing point timecodes, or the length and comprehensiveness of passages timecodes. No standards have been set as to how these choices are made and there are strengths and weaknesses of the different approaches. I hope to have the opportunity to compare notes with others using the various models, determine the trades-offs between models, establish what can and cannot be standardized, and allow digital oral history stewards to proceed with future investments in software more informed.
 

Wednesday, March 13, 2013

Decision Making as the Essence of Preservation, Curation, and Dissemination of Oral History



Escaping "The Digital Mortgage"
 

In discussions about storage, preservation, and archiving of digital video of oral histories, Doug Boyd frequently refers to the “Digital Mortgage”.  The digital mortgage is a catch-all concept meant to emphasize long-term the commitment associated with oral histories recorded on digital media. One major challenge with digital video is that equipment is used, marketed, and continues to be developed not for archival purposes but rather for production purposes. Thus oral historians and others using these media for full-length audio/video documentation must become defacto archivists, with no feasible workflow models available in academia, commercial practice, nor in film and television production.  Whereas archiving digital audio is a little easier to get a handle on, we still see professional contexts where unused raw media has transitioned only from the literal cutting room floor to a virtual cutting room floor.  Although “best practice” strategies may be recommended for video, can even well-funded collection stewards sustain these methods uniformly across holdings? Although consistency in archival quality is nice, as Boyd points out that "perfect" may be "the enemy of good enough."  Reality begs that different material within a collection be treated differently and the risk of loss should be considered on the collection and even the tape level, rather than a model that calls for a blank check into the vastly unknown and uncharted digital video future. Before signing anything, get a grip on what you have first so you can then make good choices about their value.

This is an adaptation of an abstract I drafted but never submitted anywhere, although some of the ideas emerge in our forthcoming essay in the Oral History Review, "Digital Curation through Information Cartography: A commentary on oral history in the digital age from a content management vantage point" (Lambert and Frisch, 2013)

Wednesday, February 20, 2013

Indexes as Evaluative Tools



When I began my Ph.D studies, my graduate advisor gave me a valuable piece of advice: you will not have the time or energy to pursue every good idea you have. This is not only solid life advice regarding personal time management but holds very true for "oral history in the digital age". Not every recorded oral history is destined to make it into a PBS documentary--most won't. With a good index, or sometimes just a decent inventory, one can make choices about what material is strongest and where, when, why and how it should be put to use in any of the increasing multi-media avenues available. 


Wednesday, February 13, 2013

Meaning Mapping for Ballparks, not Bulls eyes


In a number of our projects, we have had difficulty training our clients and partners NOT to get too specific by trying to imagine every possible future user while creating the controlled vocabularies and multi-dimensional indexes. (When we did this on an early project we ended up with an “out-of-control” controlled vocabulary.) The but the indexing process is not about naming things, rather sorting them into a collection of baskets of meaning we create. The structure of baskets can grow, change, expand or contract over time (an "iterative" process) and provide a lot of retrieval power and more than ample browsing power.

A misleading concept that Google reinforces in the digital age is the idea that accessing multi-media is about hitting bulls eyes or getting home runs. But one of the most underrated features of Google is not its powerful secret engine for retrieval that we will never understand, but the way it now leverages 10+ years of our “near misses”.  Google’s correlative database gives us a quick list of potential things we meant, but, like many others users, have misspelled or mis-Googled and eventually found. Even Google knows that searches are actually less about searching, and more about matching meaning to users’ desires. "Browsing" is still a mode of operating on the web and "Googling" is something different--more specific. We remember and embrace the promise of the web as a place to explore and browse...

Wednesday, December 12, 2012

Representation of recordings through Annotation




Our practice of Oral History content management, which we often refer to as digital indexing, began by questioning the assumption that recordings must be transcribed word for word before they can be used.  In the database-driven environments we work in, summary annotations are much preferred to full transcription. The work of the annotator is at the heart of digital indexing and we continue to reiterate that full transcription is always an option if the need is there and time and resources allow. Here are some random musings about annotation...

  • Annotation is about representing what is on a recording in a text format.  In this sense, it is no different than a transcription.
  • An annotation needs to describe passages of audio adequately enough to lead a user to that passage. Defining the users well may be as important or more important than the specificity with which the annotation represents the passage.
  • An annotation can be enhanced by using strategic vocabulary words within the prose of the annotation. Thus full text searches will get hits on that digital object (passage of audio or video). 
  • All annotations are subjective, and that is totally okay. It is all part of recognizing, defining, and composing toward an audience of users. Our subjectivity saves them time.
Phone: 800-554-1047 - E-mail: info@randforce.com
Web Site Copyright © 2011 The Randforce Associates, LLC