Q&A or Document-Based? The Effects of Interface Type on How Screen Reader Users Access Interconnected Documents

📄 arXiv: 2608.25382v1 📥 PDF

作者: Colleen F. Cipriano, Yichun Zhao, Miguel A. Nacenta, Kotaro Hara, Jaylee Soh

分类: cs.HC, cs.AI, cs.IR

发布日期: 2026-08-26

备注: 17 pages, 12 figures, accepted at ASSETS 2026

DOI: 10.1145/3797867.3829036


💡 一句话要点

比较问答界面与文档界面对盲人用户知识构建的影响

🎯 匹配领域: 支柱九:具身大模型 (Embodied Foundation Models)

关键词: 盲人用户 知识构建 界面设计 问答系统 文档导航 用户体验 信息访问

📋 核心要点

  1. 现有的问答界面在支持盲人用户构建知识时存在不足,难以有效整合信息。
  2. 本研究通过比较问答界面与文档界面,探讨不同界面对知识构建的影响,旨在优化用户体验。
  3. 实验结果表明,文档界面在知识构建上表现更佳,参与者在此界面中形成的心理模型更准确。

📝 摘要(中文)

盲人及低视力用户越来越多地使用大型语言模型接口来访问文档,但这些系统如何支持或阻碍他们构建互联知识尚不清楚。为了解决这一问题,研究比较了支持开放式对话的问答界面(QAI)与基于传统结构化文本导航的文档界面(DI)。通过招募16名盲人用户,使用两种界面探索两个虚构世界,数据分析显示,参与者在DI中访问了更多不同的文档,并形成了更大且更准确的心理模型。尽管如此,许多参与者仍偏好QAI,认为在QAI中探索更多并形成更好的心理模型。研究分析了这些差异的界面设计原因,并强调了使用问答界面访问信息空间的风险。

🔬 方法详解

问题定义:本研究旨在解决盲人及低视力用户在使用大型语言模型接口时,如何有效构建互联知识的问题。现有的问答界面(QAI)可能无法充分支持用户的信息整合能力。

核心思路:通过比较问答界面与传统文档界面(DI),研究探讨不同界面设计对用户知识构建的影响,旨在找出更有效的界面设计方案。

技术框架:研究招募了16名盲人用户,使用两种不同的界面探索两个虚构世界。数据收集包括交互日志、概念图、决策任务和半结构化访谈,分析用户在不同界面下的表现。

关键创新:本研究的创新之处在于系统比较了问答界面与文档界面对知识构建的影响,揭示了用户偏好与实际效果之间的差异,提供了界面设计的新视角。

关键设计:在实验中,文档界面设计强调结构化文本导航,允许用户更有效地访问和整合信息,而问答界面则侧重于开放式对话,可能导致用户对信息的误估。具体的参数设置和用户交互流程在研究中进行了详细记录。

🖼️ 关键图片

fig_0
fig_1
fig_2

📊 实验亮点

实验结果显示,参与者在文档界面中访问了更多不同的文档,并形成了更大且更准确的心理模型。相比之下,尽管在问答界面中用户的主观感受更好,但实际知识应用能力却较弱。这一发现强调了界面设计对知识构建的重要性。

🎯 应用场景

该研究的结果对设计盲人及低视力用户的文档访问界面具有重要意义,能够为未来的界面设计提供指导,提升信息获取的有效性和用户体验。潜在应用包括教育、信息检索和辅助技术等领域。

📄 摘要(原文)

Blind and low-vision (BLV) users are increasingly engaging with large language model (LLM) interfaces to access documents, but it is unclear how such systems support or hinder their ability to build interconnected knowledge. To examine this gap, we compared a Question-Answer Interface (QAI) that supports open-ended conversational inquiry, with a Document Interface (DI) based mostly on traditional structured text document navigation. We recruited 16 BLV screen reader users where they used both interfaces to explore two fictional worlds. Data from interaction logs, concept maps, decision-based tasks, and semi-structured interviews provide comparative insights into how interface design supports knowledge construction. Findings show that participants visited more distinct documents with the DI and formed larger and more correct mental models with the DI than with the QAI. They were also more able to apply knowledge they had gained. Simultaneously, many still preferred the QAI and often estimated that they had explored more, formed better mental models and applied their models better when acquiring the information with the QAI, despite this not being the case. Our analysis suggests possible interface design reasons for these differences and highlights some of the risks introduced by using question-answer interfaces to access information spaces.