School of Information Sciences

Underwood receives NEH grant to investigate consequences of error in digital libraries

Ted Underwood
Ted Underwood, Professor

Professor Ted Underwood has received a $73,122 grant from the National Endowment for the Humanities to investigate the consequences of error in digital libraries. While digital libraries represent an immense storehouse of knowledge, the texts are full of errors because of the imperfect process by which they are transcribed optically.

"It isn't unusual for five percent of the words in volumes to be mistranscribed, with the level of error much higher in some volumes," said Underwood. "Simply measuring the fraction of mistranscribed words is easy. It’s harder to know how much difference those errors make for the methods and questions that actually interest researchers. Some forms of analysis are undisturbed by high levels of error; others may be quite sensitive, especially when errors are distributed unevenly across different historical periods and genres."

Underwood will work with graduate students from the iSchool and English Department to construct parallel collections that pair each "clean" text with a realistically error-ridden version of the same book drawn from a digital library. The team will build collections of Chinese texts as well as English texts ranging from 1700 to the present, because different character sets and printing technologies produce different kinds of error. Then the team will apply a wide range of data-mining methods to both the clean and error-ridden collections and measure the distortion produced by transcription error and other common sources of noise. The project will provide tools that help other researchers estimate the level of uncertainty in their own conclusions.

"No data is perfect. There's always some kind of error. The question is whether the error is of a kind and magnitude likely to matter for a particular question," he said.

Underwood is a professor in the iSchool and also holds an appointment with the Department of English in the College of Liberal Arts and Sciences. He has authored three books about literary history, including Distant Horizons (The University of Chicago Press Books, 2019), Why Literary Periods Mattered: Historical Contrast and the Prestige of English Studies (Stanford University Press, 2013), and The Work of the Sun: Literature, Science and Political Economy 1760-1860 (New York: Palgrave, 2005). His articles have appeared in PMLA, Representations, MLQ, and Cultural Analytics. Underwood earned his PhD in English from Cornell University.

Updated on
Backto the news archive

Related News

Bashir group presents work at PEPR 2026

PhD students Ramazan Yener, Eryue Xu, and Mubarak Raji presented their research this week at the 2026 USENIX Conference on Privacy Engineering Practice and Respect (PEPR) in Santa Clara, California. PEPR is focused on designing and building products and systems with privacy and respect for their users and the societies in which they operate. The students received USENIX grants covering their conference registration and providing travel support to attend the conference. 

Bashir group PEPR 2026

2025 Downs Intellectual Freedom Award given to Nicole A. Cooke

Nicole A. Cooke has been named the 2025 recipient of the Downs Intellectual Freedom Award for her advocacy, groundbreaking research, and dedication to diversity, equity, and inclusion within the field of library and information science. Cooke is the Augusta Baker Endowed Chair and professor in the College of Information and Communications at the University of South Carolina.

Nicole Cooke

iSchool researchers to present work at CVPR Conference

Assistant Professors Ismini Lourentzou and Yaoyao Liu, along with students from their labs, will present their research at the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), held in Denver, Colorado, from June 3–7. CVPR is the flagship annual meeting of IEEE/CVF and PAMI-TC, where researchers present their latest advances in computer vision, pattern recognition, machine learning, robotics, and artificial intelligence, both in theory and practice. 

iSchool alumni named 2026 Movers & Shakers

Two iSchool alumni are included in Library Journal's 2026 class of Movers & Shakers, an annual list that recognizes 50 professionals who are moving the library field as a profession. Leah T. Dudak (MSLIS '17) was honored in the Advocates category and Mariella Colon (MSLIS '07) was honored in the Community Builders category. 

iSchool researchers to present at ChLA 2026

iSchool faculty and staff will present their research at the Children's Literature Association (ChLA) annual conference, which will be held from May 28-30 in Pittsburgh, Pennsylvania. The theme of this year's conference is "Neighbors and Neighborhoods in Children's Literature, Media, and Culture."

School of Information Sciences

501 E. Daniel St.

MC-493

Champaign, IL

61820-6211

Voice: (217) 333-3280

Email: ischool@illinois.edu

Back to top