Steering LLM Summarization with Visual Workspaces for Sensemaking

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Tang, Xuxin, Krokos, Eric, Liu, Can, Davidson, Kylie, Whitley, Kirsten, Ramakrishnan, Naren, North, Chris
Format: Preprint
Veröffentlicht: 2024
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866908061338173440
author Tang, Xuxin
Krokos, Eric
Liu, Can
Davidson, Kylie
Whitley, Kirsten
Ramakrishnan, Naren
North, Chris
author_facet Tang, Xuxin
Krokos, Eric
Liu, Can
Davidson, Kylie
Whitley, Kirsten
Ramakrishnan, Naren
North, Chris
contents Large Language Models (LLMs) have been widely applied in summarization due to their speedy and high-quality text generation. Summarization for sensemaking involves information compression and insight extraction. Human guidance in sensemaking tasks can prioritize and cluster relevant information for LLMs. However, users must translate their cognitive thinking into natural language to communicate with LLMs. Can we use more readable and operable visual representations to guide the summarization process for sensemaking? Therefore, we propose introducing an intermediate step--a schematic visual workspace for human sensemaking--before the LLM generation to steer and refine the summarization process. We conduct a series of proof-of-concept experiments to investigate the potential for enhancing the summarization by GPT-4 through visual workspaces. Leveraging a textual sensemaking dataset with a ground truth summary, we evaluate the impact of a human-generated visual workspace on LLM-generated summarization of the dataset and assess the effectiveness of space-steered summarization. We categorize several types of extractable information from typical human workspaces that can be injected into engineered prompts to steer the LLM summarization. The results demonstrate how such workspaces can help align an LLM with the ground truth, leading to more accurate summarization results than without the workspaces.
format Preprint
id arxiv_https___arxiv_org_abs_2409_17289
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Steering LLM Summarization with Visual Workspaces for Sensemaking
Tang, Xuxin
Krokos, Eric
Liu, Can
Davidson, Kylie
Whitley, Kirsten
Ramakrishnan, Naren
North, Chris
Human-Computer Interaction
Large Language Models (LLMs) have been widely applied in summarization due to their speedy and high-quality text generation. Summarization for sensemaking involves information compression and insight extraction. Human guidance in sensemaking tasks can prioritize and cluster relevant information for LLMs. However, users must translate their cognitive thinking into natural language to communicate with LLMs. Can we use more readable and operable visual representations to guide the summarization process for sensemaking? Therefore, we propose introducing an intermediate step--a schematic visual workspace for human sensemaking--before the LLM generation to steer and refine the summarization process. We conduct a series of proof-of-concept experiments to investigate the potential for enhancing the summarization by GPT-4 through visual workspaces. Leveraging a textual sensemaking dataset with a ground truth summary, we evaluate the impact of a human-generated visual workspace on LLM-generated summarization of the dataset and assess the effectiveness of space-steered summarization. We categorize several types of extractable information from typical human workspaces that can be injected into engineered prompts to steer the LLM summarization. The results demonstrate how such workspaces can help align an LLM with the ground truth, leading to more accurate summarization results than without the workspaces.
title Steering LLM Summarization with Visual Workspaces for Sensemaking
topic Human-Computer Interaction
url https://arxiv.org/abs/2409.17289