Beyond Pixels: Introspective and Interactive Grounding for Visualization Agents
Fuente:
arXiv
Guardado en:
| Autores principales: | Lu, Yiyang, Shin, Woong, Karimi, Ahmad Maroof, Wang, Feiyi, Ren, Jie, Smirni, Evgenia |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Profiling and Modeling of Power Characteristics of Leadership-Scale HPC System Workloads
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024)
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024)
Exploring the Frontiers of Energy Efficiency using Power Management at System Scale
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024)
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024)
EPIC: Generative AI Platform for Accelerating HPC Operational Data Analytics
por: Karimi, Ahmad Maroof, et al.
Publicado: (2025)
por: Karimi, Ahmad Maroof, et al.
Publicado: (2025)
Chatting with Images for Introspective Visual Thinking
por: Wu, Junfei, et al.
Publicado: (2026)
por: Wu, Junfei, et al.
Publicado: (2026)
Beyond Introspection: Reinforcing Thinking via Externalist Behavioral Feedback
por: Yang, Diji, et al.
Publicado: (2024)
por: Yang, Diji, et al.
Publicado: (2024)
PixelPrune: Pixel-Level Adaptive Visual Token Reduction via Predictive Coding
por: Wang, Nan, et al.
Publicado: (2026)
por: Wang, Nan, et al.
Publicado: (2026)
From Imitation to Introspection: Probing Self-Consciousness in Language Models
por: Chen, Sirui, et al.
Publicado: (2024)
por: Chen, Sirui, et al.
Publicado: (2024)
I3: Intent-Introspective Retrieval Conditioned on Instructions
por: Pan, Kaihang, et al.
Publicado: (2023)
por: Pan, Kaihang, et al.
Publicado: (2023)
STAIR: Improving Safety Alignment with Introspective Reasoning
por: Zhang, Yichi, et al.
Publicado: (2025)
por: Zhang, Yichi, et al.
Publicado: (2025)
AI-Gram: When Visual Agents Interact in a Social Network
por: Shin, Andrew
Publicado: (2026)
por: Shin, Andrew
Publicado: (2026)
Beyond Pixels: Exploring DOM Downsampling for LLM-Based Web Agents
por: Schiepanski, Thassilo M., et al.
Publicado: (2025)
por: Schiepanski, Thassilo M., et al.
Publicado: (2025)
Recursive Introspection: Teaching Language Model Agents How to Self-Improve
por: Qu, Yuxiao, et al.
Publicado: (2024)
por: Qu, Yuxiao, et al.
Publicado: (2024)
Emergent Introspection in AI is Content-Agnostic
por: Lederman, Harvey, et al.
Publicado: (2026)
por: Lederman, Harvey, et al.
Publicado: (2026)
Emergent Introspective Awareness in Large Language Models
por: Lindsey, Jack
Publicado: (2026)
por: Lindsey, Jack
Publicado: (2026)
Privileged Self-Access Matters for Introspection in AI
por: Song, Siyuan, et al.
Publicado: (2025)
por: Song, Siyuan, et al.
Publicado: (2025)
VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents
por: Hu, Jiliang, et al.
Publicado: (2025)
por: Hu, Jiliang, et al.
Publicado: (2025)
Grounded Language Agent for Product Search via Intelligent Web Interactions
por: Fereidouni, Moghis, et al.
Publicado: (2024)
por: Fereidouni, Moghis, et al.
Publicado: (2024)
Language Models Fail to Introspect About Their Knowledge of Language
por: Song, Siyuan, et al.
Publicado: (2025)
por: Song, Siyuan, et al.
Publicado: (2025)
Decoding Memories: An Efficient Pipeline for Self-Consistency Hallucination Detection
por: Gao, Weizhi, et al.
Publicado: (2025)
por: Gao, Weizhi, et al.
Publicado: (2025)
Weakly-Supervised 3D Visual Grounding based on Visual Language Alignment
por: Xu, Xiaoxu, et al.
Publicado: (2023)
por: Xu, Xiaoxu, et al.
Publicado: (2023)
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning
por: Wang, Haozhe, et al.
Publicado: (2025)
por: Wang, Haozhe, et al.
Publicado: (2025)
PatchBoard: Schema-Grounded State Mutation for Reliable and Auditable LLM Multi-Agent Collaboration
por: Zhang, Shuyu, et al.
Publicado: (2026)
por: Zhang, Shuyu, et al.
Publicado: (2026)
ConflictBench: Evaluating Human-AI Conflict via Interactive and Visually Grounded Environments
por: Zhao, Weixiang, et al.
Publicado: (2026)
por: Zhao, Weixiang, et al.
Publicado: (2026)
Does It Make Sense to Speak of Introspection in Large Language Models?
por: Comsa, Iulia M., et al.
Publicado: (2025)
por: Comsa, Iulia M., et al.
Publicado: (2025)
Scaling Agentic Capabilities via Grounded Interaction Synthesis
por: Shi, Wenhang, et al.
Publicado: (2026)
por: Shi, Wenhang, et al.
Publicado: (2026)
ValueGround: Evaluating Culture-Conditioned Visual Value Grounding in MLLMs
por: Wang, Zhipin, et al.
Publicado: (2026)
por: Wang, Zhipin, et al.
Publicado: (2026)
AgentReview: Exploring Peer Review Dynamics with LLM Agents
por: Jin, Yiqiao, et al.
Publicado: (2024)
por: Jin, Yiqiao, et al.
Publicado: (2024)
Beyond Words: Evaluating and Bridging Epistemic Divergence in User-Agent Interaction via Theory of Mind
por: Ruan, Minyuan, et al.
Publicado: (2026)
por: Ruan, Minyuan, et al.
Publicado: (2026)
Grounding Language Models for Visual Entity Recognition
por: Xiao, Zilin, et al.
Publicado: (2024)
por: Xiao, Zilin, et al.
Publicado: (2024)
Aging Up AAC: An Introspection on Augmentative and Alternative Communication Applications for Autistic Adults
por: Martin, Lara J., et al.
Publicado: (2024)
por: Martin, Lara J., et al.
Publicado: (2024)
Beyond One Path: Evaluating and Enhancing Divergent Thinking in Interactive LLM Agents
por: Park, Jihyeong, et al.
Publicado: (2026)
por: Park, Jihyeong, et al.
Publicado: (2026)
Retrieve, Integrate, and Synthesize: Spatial-Semantic Grounded Latent Visual Reasoning
por: Cui, Jin, et al.
Publicado: (2026)
por: Cui, Jin, et al.
Publicado: (2026)
Pixology: Probing the Linguistic and Visual Capabilities of Pixel-based Language Models
por: Tatariya, Kushal, et al.
Publicado: (2024)
por: Tatariya, Kushal, et al.
Publicado: (2024)
GroundAct: Can LLM Agents Ground Actions in Environmental States?
por: Wang, Zixuan, et al.
Publicado: (2025)
por: Wang, Zixuan, et al.
Publicado: (2025)
VisualTrap: A Stealthy Backdoor Attack on GUI Agents via Visual Grounding Manipulation
por: Ye, Ziang, et al.
Publicado: (2025)
por: Ye, Ziang, et al.
Publicado: (2025)
Looking Inward: Language Models Can Learn About Themselves by Introspection
por: Binder, Felix J, et al.
Publicado: (2024)
por: Binder, Felix J, et al.
Publicado: (2024)
Learning to Use Tools via Cooperative and Interactive Agents
por: Shi, Zhengliang, et al.
Publicado: (2024)
por: Shi, Zhengliang, et al.
Publicado: (2024)
From Pixels to Tokens: Byte-Pair Encoding on Quantized Visual Modalities
por: Zhang, Wanpeng, et al.
Publicado: (2024)
por: Zhang, Wanpeng, et al.
Publicado: (2024)
Zero-Overhead Introspection for Adaptive Test-Time Compute
por: Manvi, Rohin, et al.
Publicado: (2025)
por: Manvi, Rohin, et al.
Publicado: (2025)
I-MCTS: Enhancing Agentic AutoML via Introspective Monte Carlo Tree Search
por: Liang, Zujie, et al.
Publicado: (2025)
por: Liang, Zujie, et al.
Publicado: (2025)
Ejemplares similares
-
Profiling and Modeling of Power Characteristics of Leadership-Scale HPC System Workloads
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024) -
Exploring the Frontiers of Energy Efficiency using Power Management at System Scale
por: Karimi, Ahmad Maroof, et al.
Publicado: (2024) -
EPIC: Generative AI Platform for Accelerating HPC Operational Data Analytics
por: Karimi, Ahmad Maroof, et al.
Publicado: (2025) -
Chatting with Images for Introspective Visual Thinking
por: Wu, Junfei, et al.
Publicado: (2026) -
Beyond Introspection: Reinforcing Thinking via Externalist Behavioral Feedback
por: Yang, Diji, et al.
Publicado: (2024)