Downie to discuss HTRC findings at Harvard Library

Stephen Downie
J. Stephen Downie, Professor, Associate Dean for Research, and Co-Director of the HathiTrust Research Center

Professor and Associate Dean for Research J. Stephen Downie will present his recent work with the HathiTrust Research Center (HTRC) on April 30 at Harvard Library. Downie is codirector of HTRC, a collaboration between the University of Illinois, Indiana University, and the HathiTrust to enable advanced computational access to text found in the HathiTrust (HT) Digital Library.

His talk, "Creating Universal Open Access to Closed Textual Data at Scale: Use Cases from the HathiTrust Research Center," will discuss how the HTRC is creating a set of non-consumptive research services to make HT Digital Library volumes that are under copyright restrictions more open and useful to scholars.

"The creation and publication of the HTRC 'Extracted Features' (EF) dataset provides unigram counts and Part-of-Speech (POS) information for each of the 5.6 billion pages in the HT Digital Library," explained Downie. "In my talk, I will introduce two uses cases that leverage the EF dataset: the 'HathiTrust + Bookworm' visualization and analysis tool; and the Workset Building environment developed to provide researchers fine-grained access to the entire HT collection (both public domain and in-copyright) via the EF dataset."

Downie leads the HathiTrust + Bookworm text analysis project, which is creating tools to visualize the evolution of term usage over time. He also is the principal investigator on the Workset Creation for Scholarly Analysis + Data Capsules project, which integrates workset models and tools, and he represents the HTRC on the Novel(TM) text mining project as well as the Single Interface for Music Score Searching and Analysis project. All of these projects strive to provide large-scale analytic access to copyright-restricted cultural data.

Research Areas:
Updated on
Backto the news archive

Related News

Knox recognized for public engagement

Associate Professor Emily Knox has been selected as the recipient of the Campus Excellence in Public Engagement Emerging Award. She will be honored on May 28 at a special event hosted by the Office of Public Engagement. 

Emily Knox

Schneider selected as 2024-2025 Harvard Radcliffe Institute Fellow

Associate Professor Jodi Schneider has been selected as a 2024-2025 fellow of the Harvard Radcliffe Institute, an institute of Harvard University that fosters interdisciplinary research across the humanities, sciences, social sciences, arts, and professions.

Jodi Schneider

iSchool researchers to present at ACM Web Conference

Members of Associate Professor Dong Wang's research group, the Social Sensing and Intelligence Lab, will present their research at the Web Conference 2024, which will be held from May 13-17 in Singapore. The Web Conference is the premier venue to present and discuss progress in research, development, standards, and applications of topics related to the Web.

iSchool researchers to present at CHI 2024

iSchool faculty and students will present their research at the ACM Conference on Human Factors in Computing Systems (CHI 2024), which will be held from May 11-16 in Honolulu, Hawaii. The conference, considered the most prestigious in the field of Human-Computer Interaction, attracts researchers and practitioners from around the globe. The theme for CHI 2024 is "Surfing the World."

CHI 2024

iSchool researchers present at inaugural ASIS&T symposium

iSchool researchers will present their work at the Association for Information Science & Technology (ASIS&T) Midwest Chapter Spring Symposium on April 26. The inaugural symposium will include talks by seventeen researchers from ten institutions across the Midwest region.