AVA: Towards Autonomous Visualization Agents through Visual Perception-Driven Decision-Making

Liu, Shusen; Miao, Haichao; Li, Zhimin; Olson, Matthew; Pascucci, Valerio; Bremer, Peer-Timo

AVA: Towards Autonomous Visualization Agents through Visual Perception-Driven Decision-Making

dc.contributor.author	Liu, Shusen	en_US
dc.contributor.author	Miao, Haichao	en_US
dc.contributor.author	Li, Zhimin	en_US
dc.contributor.author	Olson, Matthew	en_US
dc.contributor.author	Pascucci, Valerio	en_US
dc.contributor.author	Bremer, Peer-Timo	en_US
dc.contributor.editor	Aigner, Wolfgang	en_US
dc.contributor.editor	Archambault, Daniel	en_US
dc.contributor.editor	Bujack, Roxana	en_US
dc.date.accessioned	2024-05-21T08:18:38Z
dc.date.available	2024-05-21T08:18:38Z
dc.date.issued	2024
dc.description.abstract	With recent advances in multi-modal foundation models, the previously text-only large language models (LLM) have evolved to incorporate visual input, opening up unprecedented opportunities for various applications in visualization. Compared to existing work on LLM-based visualization works that generate and control visualization with textual input and output only, the proposed approach explores the utilization of the visual processing ability of multi-modal LLMs to develop Autonomous Visualization Agents (AVAs) that can evaluate the generated visualization and iterate on the result to accomplish user-defined objectives defined through natural language. We propose the first framework for the design of AVAs and present several usage scenarios intended to demonstrate the general applicability of the proposed paradigm. Our preliminary exploration and proof-of-concept agents suggest that this approach can be widely applicable whenever the choices of appropriate visualization parameters require the interpretation of previous visual output. Our study indicates that AVAs represent a general paradigm for designing intelligent visualization systems that can achieve high-level visualization goals, which pave the way for developing expert-level visualization agents in the future.	en_US
dc.description.number	3
dc.description.sectionheaders	Workflows and Decision Making
dc.description.seriesinformation	Computer Graphics Forum
dc.description.volume	43
dc.identifier.doi	10.1111/cgf.15093
dc.identifier.issn	1467-8659
dc.identifier.pages	12 pages
dc.identifier.uri	https://doi.org/10.1111/cgf.15093
dc.identifier.uri	https://diglib.eg.org/handle/10.1111/cgf15093
dc.publisher	The Eurographics Association and John Wiley & Sons Ltd.	en_US
dc.title	AVA: Towards Autonomous Visualization Agents through Visual Perception-Driven Decision-Making	en_US

Files

Original bundle

Now showing 1 - 2 of 2

Name:: v43i3_18_cgf15093.pdf
Size:: 17.09 MB
Format:: Adobe Portable Document Format

Download

Name:: 1153-i7.pdf
Size:: 14.85 MB
Format:: Adobe Portable Document Format

Download

Collections

43-Issue 3