The Semantic Reader Project: Augmenting Scholarly Documents through AI-Powered Interactive Reading Interfaces
Author(s) • • • • • • • • •
Lo, Kyle
Chang, Joseph
Head, Andrew
Bragg, Jonathan
Zhang, Amy
Trier, Cassidy
Anastasiades, Chloe
August, Tal
Authur, Russell
Bragg, Danielle
Date Issued
October 1, 2024
Journal
Communications of the ACM
Publisher
ACM
Citation
Lo, Kyle, Chang, Joseph, Head, Andrew, Bragg, Jonathan, Zhang, Amy et al. 2024. "The Semantic Reader Project: Augmenting Scholarly Documents through AI-Powered Interactive Reading Interfaces." Communications of the ACM, 67 (10).
Version
Final published version
Abstract
Scholarly publications are key to the transfer of knowledge from scholars to others. However, research papers are information-dense, and as the volume of the scientific literature grows, the greater the need for new technology to support scholars. In contrast to the process of finding papers, which has been transformed by Internet technology, the experience of reading research papers has changed little in decades. For instance, the PDF format for sharing papers remains widely used due to its portability but has significant downsides, inter alia, static content and poor accessibility for low-vision readers. This paper explores the question "Can recent advances in AI and HCI power intelligent, interactive, and accessible reading interfaces, even for legacy PDFs?" We describe the Semantic Reader Project, a collaborative effort across multiple institutions to explore automatic creation of dynamic reading interfaces for research papers. Through this project, we've developed a collection of novel reading interfaces and evaluated them with study participants and real-world users to show improved reading experiences for scholars. We've also released a production research paper reading interface that will continuously incorporate novel features from our research as they mature. We structure this paper around five key opportunities for AI assistance in scholarly reading---discovery, efficiency, comprehension, synthesis, and accessibility---and present an overview of our progress and discuss remaining open challenges.
Terms of Use
Creative Commons Attribution
Persistent DSpace Link
DOI of Published Version
https://doi.org/10.1145/3659096