VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization

Sam Yu-Te LeeJingya ChenAlbert CalzarettoRichard LeeSamir PassiAlice FerngMihaela Vorvoreanu

article2025arXiv1 citations

Demonstrates that integrating document visualization with topic modeling allows knowledge workers to inspect and adjust retrieved context in enterprise chatbots, improving answer relevance and user prompting strategies.

Listen

Enterprise artificial intelligence chatbots promise to streamline knowledge work by synthesizing vast corporate data archives, yet they frequently generate answers that are plausible but misaligned with actual user intent. This creates a critical operational risk known as overreliance, where users uncritically accept flawed recommendations that can silently degrade decision-making, increase compliance risk, and harm performance. Conventional conversational interfaces leave underlying retrieval processes obscured, forcing users to rely on trial-and-error prompting without providing transparent mechanisms to inspect or steer the contextual data feeding the model.

The article evaluates whether visual context engineering—giving knowledge workers direct visual oversight and control over retrieved data—fosters appropriate reliance on enterprise chatbots. It demonstrates this approach through a research prototype, VizCopilot, which couples a conversational assistant with an interactive treemap visualization.

To assess this paradigm, the authors conducted a qualitative Research-through-Design study comparing VizCopilot against a baseline pure-text chat interface. Fourteen participants with prior enterprise chatbot experience completed realistic information synthesis tasks. The testing leveraged a synthetic corporate corpus representing approximately 1,000 employees and 10,000 messy enterprise records, complete with realistic metadata conflicts, duplicate items, and unstructured content.

The investigation produced four primary findings. First, interactive visual scaffolding enabled users to rapidly detect context misalignment, such as missing topics or extraneous data, without manually reviewing every file. Second, direct manipulation allowed users to successfully correct errors: for instance, 10 out of 14 participants detected and resolved a subtle error where the chatbot conflated two distinct employees sharing the same name by inspecting the visual file view. Third, visual context substantially improved user agency and efficiency; participants required fewer follow-up prompts to achieve target results and naturally adapted their queries using highlighted keywords. Finally, participants showed persistent skepticism toward automated subtopic text summaries, consistently preferring direct access to raw evidence for critical verification.

These findings indicate that integrating interactive visual structures can meaningfully augment human oversight and reduce overreliance without creating excessive cognitive strain. By exposing intermediate retrieval steps, the interface counters anthropomorphic assumptions about artificial intelligence, encouraging users to treat the tool as an inspectable search-and-synthesis system. This balance between automation and direct manipulation offers a viable pathway toward meeting emerging regulatory requirements for human oversight in high-risk automated systems.

Organizations developing or deploying enterprise chatbots should consider incorporating group-level visual context controls alongside conversational prompts to improve retrieval alignment. System designers must also provide explicit uncertainty indicators to trigger manual verification on deceptively simple queries and improve tools for rapid raw-document inspection, such as keyword highlighting. Moving forward, teams should pilot these interfaces across larger real-world data environments to evaluate technical scalability and refine onboarding workflows that ease initial interface complexity.

The conclusions should be interpreted within the boundaries of an exploratory qualitative study using a synthetic dataset and brief participant sessions. While long-term longitudinal studies are needed before scaling across production enterprise suites, the article provides moderate to high confidence that visual context engineering is a superior design direction for maintaining human agency and reliable system oversight.

arXiv: 2510.11954
Cover for VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization

Abstract

Enterprise chatbots show promise in supporting knowledge workers in information synthesis tasks by retrieving context from large, heterogeneous databases before generating answers. However, when the retrieved context misaligns with user intentions, the chatbot often produces "irrelevantly right" responses that provide little value. In this work, we introduce VizCopilot, a prototype that incorporates visualization techniques to actively involve end-users in context alignment. By combining topic modeling with document visualization, VizCopilot enables human oversight and modification of retrieved context while keeping cognitive overhead manageable. We used VizCopilot as a design probe in a Research-through-Design study to evaluate the role of visualization in context alignment and to surface future design opportunities. Our findings show that visualization not only helps users detect and correct misaligned context but also encourages them to adapt their prompting strategies, enabling the system to retrieve more relevant context from the outset. At the same time, the study reveals limitations in verification support regarding close-reading and trust in AI summaries. We outline future directions for visualization-enhanced chatbots, focusing on personalization, proactivity, and sustainable human-AI collaboration.

Citation

MLA
Lee, S. Y.-T., et al. “VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization”. arXiv, 2025, http://arxiv.org/abs/2510.11954v2.
APA
Lee, S. Y.-T., Chen, J., Calzaretto, A., Lee, R., Passi, S., Ferng, A., & Vorvoreanu, M. (2025). VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization. arXiv. http://arxiv.org/abs/2510.11954v2
Chicago
Lee, S. Y.-T., J. Chen, A. Calzaretto, et al. 2025. “VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization”. arXiv. http://arxiv.org/abs/2510.11954v2.
Harvard
Lee, S.Y.-T. et al. (2025) “VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization”, arXiv [Preprint]. Available at: http://arxiv.org/abs/2510.11954v2.
Vancouver
1. Lee SY-T, Chen J, Calzaretto A, Lee R, Passi S, Ferng A, Vorvoreanu M (2025) VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization. arXiv

BibTeX

@article{lee2025vizcopilot,
  title = {VizCopilot: Fostering Appropriate Reliance on Enterprise Chatbots with Context Visualization},
  author = {Lee, Sam Yu-Te and Chen, Jingya and Calzaretto, Albert and Lee, Richard and Passi, Samir and Ferng, Alice and Vorvoreanu, Mihaela},
  year = {2025},
  journal = {arXiv},
  url = {http://arxiv.org/abs/2510.11954v2},
  eprint = {2510.11954}
}
Metadata:arXiv

Access the Paper

This paper is available from its original source. Click below to access the PDF.

Open PDF
License: https://creativecommons.org/licenses/by/4.0/