Beyond the Job Title: Research Data Librarian
A library tech job interview with Kaylee Alexander-Leunissen
Posted on in Job Profiles
Posted on December 23, 2024 in Job Profiles
Authors:
Eric Lease Morgan
Core to LibTech Insights’s mission is demystifying the broad and dynamic field of tech librarianship in higher ed. In this series, we interview a librarian every month to learn a little more about their position. As library tech jobs proliferate, they sometimes come with unfamiliar, jargony, or intimidating titles. We want to go beyond the title and look at the responsibilities, skills, and joys that make up the job. We hope this series will increase your knowledge of library tech jobs and skills and offer you greater insight into the working lives of your colleagues.
For this installment, we spoke to Eric Lease Morgan to learn more about his job as a Digital Initiatives Librarian. Check out our archive of job profiles. 💫
I believe my official title is Digital Initiatives Librarian, but I just call myself a librarian. I have three responsibilities:
1. Teach/facilitate workshops on the topics of text mining and natural language processing. Over the past number of years, I have taught/facilitated about three one-hour workshops per week during the academic year. These workshops include introduction to natural language processing (NLP), Python and NLP, concordancing, topic modeling, how to use the Distant Reader, how to make a book, and how to write in a book. By now, I have easily instructed hundreds of students per year.
2. Work collaboratively on projects with undergraduates, graduate students, and faculty. All these projects include components of text mining. Example research questions have included:
In all of these cases, I first amass and curate large collections of text, and these things are really datasets. (Think “collections as data.”) I then model—analyze—the datasets to address the research questions. For example, I have harvested 750 books from HathiTrust, and each book was 750 pages long. I have used the Nexis Uni API to download 48,000 newspaper articles. I have worked with a graduate student to curate a collection of 600 Victorian novels. In these ways, I practice every aspect of librarianship: collections, acquisitions, cataloging, preservation, and dissemination. In many of these cases, the results of the research are journal articles, and I appear as a coauthor.
3. Investigate how computer technology can be exploited to improve the processes of librarianship. I have been doing this for the whole of my 40-year career. Examples have included: information retrieval, personalization, usability, and artificial intelligence (AI). These things have been manifested as the Mr. Serials Process, the Alex Catalogue of Electronic texts, and MyLibrary. I wrote my first AI program in 1992, but at that time, AI systems were called “expert systems.” Recently, I have gotten a proof-of-concept grant from Amazon to explore how generative AI can be used in libraries. During the pandemic, I was awarded close to $.75 million dollars of services from Microsoft to collect, curate, and analyze 1 million scholarly journal articles on the topic of COVID-19.
🔥 Stay up-to-date with LibTech Insights by signing up for our free newsletter. Just one weekly email with our new blog posts, top tech news stories, and other bonus content. Check out some posts from our archive:
For the past five or six years, I have been developing an ecosystem called the Distant Reader. Given an almost arbitrary amount of content of almost any type, the Reader creates datasets that can be analyzed in a myriad of ways. To demonstrate the Reader’s functionality, I have created a collection of 3,000 such data sets—affectionately called “study carrels.” Moreover, to make it easier for people to create study carrels, I have created a collection of .7 million items with two different interfaces: a traditional catalog and a less traditional but more functional index. Academics are expected to read a lot. The Reader facilitates the process; the Reader makes it easy to get one’s head around dozens of books or hundreds of articles. Believe it or not, the hard part is actually getting the content, not the analysis.
In a similar vein, I have helped other librarians use computers to do library better; more or less, I started and fostered the Code4Lib community. It began as a mailing list in 2004, and it has matured to include an annual conference, a refereed journal, and a number of regional Code4Lib communities. Currently, the mailing has about 3,900 subscribers. Larger than LITA used to be?
In 1984, I was on the lending side of interlibrary loan at Drexel University, and I wanted to be a reference librarian. My boss said, “You will have to write an annual report; you will have to count and tabulate all of those little pieces of paper.” Well, I hated counting and tabulating little pieces of paper, so I wrote a program that created my annual report daily. It was then that I learned how computers could be exploited in libraries. It changed my trajectory. I now wanted to be a “systems librarian.” Later, I went on to be a medical librarian, and I got grants from Apple Computer and the National Library of Medicine. That was when I wrote my first AI program. I outgrew that job, and then I worked at the NC State Libraries, where I was one of the first 10,000 people in the world to create a website. Really. I outgrew that job too, became webmaster here at the University of Notre Dame, and after 20 years, I now work in a digital scholarship center doing the things outlined above.
As a humanist with a liberal arts education, I have always been interested in “great ideas”: truth, beauty, honor, justice, love, art, science, philosophy, religion, government, history, etc. I became a librarian as a way to be immersed in a profession where these sorts of ideas can be actively investigated and explored. I now have millions of items in my collections on these great ideas, and these items are measured in the multibillions of words. Moreover, I have both access to large computers—computers the size of Walmart—as well as the skills to use them efficiently. As I head toward retirement, I see my investigations continuing. The pursuit of truth and beauty never ends.
There are a few things, not listed in priority order:
A library tech job interview with Kaylee Alexander-Leunissen
Posted on in Job Profiles
A library tech job interview with Michelle Bowers
Posted on in Job Profiles
A library tech job interview with Whitney Christopher
Posted on in Job Profiles
A library tech job interview with Axa Liauw
Posted on in Job Profiles