TractorBeam suggests that systems that facilitate exploratory research on individual documents may lead to verifiable sensemaking for users and complement tools that work across broader corpora.
Abstract
Language model-based systems which allow asking questions of documents have become popular tools for sensemaking. Despite their implied capability, these systems still suffer from issues of factuality and provenance, while encouraging confirmatory, rather than exploratory, research. We present TractorBeam, a browser extension-based mixed-initiative system that uses collaborative annotation as an interface metaphor for sensemaking, re-framing language model (LM) outputs as suggested highlights in a process that we call collaborative machine annotation. This metaphor allows us to present LM results in-context on PDF documents, directly addressing concerns of provenance and factuality, while allowing users to iteratively construct mental schemas and queries for language models directly in the context of a document. In a preliminary user study, all of our participants felt that TractorBeam enabled them evaluate and iteratively improve the model's reflection of their intended highlighting, and several found suggestions that made them reconsider their original schema. TractorBeam suggests that systems that facilitate exploratory research on individual documents may lead to verifiable sensemaking for users and complement tools that work across broader corpora.
Recent advances in retrieval-augmented generation (RAG) and large language models (LLMs) enable researchers to integrate AI into scientific workflows. However, using proprietary commercial AI systems raises concerns about transparency, reproducibility and privacy, which are essential for scientific practices. To this e...
J. Stark, S. Saikrishnan, Vikram Seenivasan et al.· 1 citation
For decades, search and recommendation systems have been optimized as distinct components within large-scale discovery platforms. The rise of generative AI is beginning to blur this boundary. At Spotify, we are exploring how large language models can evolve from tools that retrieve content into systems that reason over...
Paul N. Bennett· Proceedings of the 32nd ACM...· 0 citations
Travel literature is a unique form of hypertext that extends beyond its medium into real-world physical exploration. While conventional computational methods can easily extract surface-level ontological entities (e.g., locations, dates), the deeper epistemic and judgmental subtext that guides the discovery and evaluati...
Virginia Orlando, Alessandro Adamou, Alessio Antonini· Proceedings of the 37th ACM...· 0 citations
Archives are essential vehicles for both historical research and memory practices like genealogy. The sheer volume of documents produced via digital infrastructures, however, has posed a challenge for traditional archival processes that rely on human judgments. Some archivists have responded to this challenge by en...
Legislative knowledge evolves as an intricate hypertext in which documents are interconnected through complex, often implicit relationships. In this paper, we introduce ReSB2, a framework for retrieving and linking similar legislative bills that supports human–machine collaboration and helps reduce redundancy in the la...
Lucas G. L. Costa, Átila Souza, Elves Rodrigues et al.· Proceedings of the 37th ACM...· 0 citations
The integration of Large Language Models is transforming recommender systems, offering unprecedented capabilities for complex reasoning and natural language generation. However, their propensity to generate hallucinations (incorrect or invented information) compromises reliability and user trust, limiting their adoptio...
Andrés Felipe Solis Pino, N. Duque-Méndez, Pablo H. Ruiz et al.· Applied Sciences· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.