MAPWise: Evaluating Vision-Language Models for Advanced Map Queries
Fuente:
arXiv
Salvato in:
| Autori principali: | Mukhopadhyay, Srija, Rajgaria, Abhishek, Khatiwada, Prerana, Gupta, Vivek, Roth, Dan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
PRAISE: Enhancing Product Descriptions with LLM-Driven Structured Insights
di: Qidwai, Adnan, et al.
Pubblicazione: (2025)
di: Qidwai, Adnan, et al.
Pubblicazione: (2025)
Unraveling the Truth: Do VLMs really Understand Charts? A Deep Dive into Consistency and Robustness
di: Mukhopadhyay, Srija, et al.
Pubblicazione: (2024)
di: Mukhopadhyay, Srija, et al.
Pubblicazione: (2024)
ClickAIXR: On-Device Multimodal Vision-Language Interaction with Real-World Objects in Extended Reality
di: Khan, Dawar, et al.
Pubblicazione: (2026)
di: Khan, Dawar, et al.
Pubblicazione: (2026)
MM-Conv: A Multi-modal Conversational Dataset for Virtual Humans
di: Deichler, Anna, et al.
Pubblicazione: (2024)
di: Deichler, Anna, et al.
Pubblicazione: (2024)
Advancements and limitations of LLMs in replicating human color-word associations
di: Fukushima, Makoto, et al.
Pubblicazione: (2024)
di: Fukushima, Makoto, et al.
Pubblicazione: (2024)
Towards Interactive Intelligence for Digital Humans
di: Cai, Yiyi, et al.
Pubblicazione: (2025)
di: Cai, Yiyi, et al.
Pubblicazione: (2025)
Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models
di: Chen, Tianrun, et al.
Pubblicazione: (2024)
di: Chen, Tianrun, et al.
Pubblicazione: (2024)
Shadowless Projection Mapping for Tabletop Workspaces with Synthetic Aperture Projector
di: Okamoto, Takahiro, et al.
Pubblicazione: (2026)
di: Okamoto, Takahiro, et al.
Pubblicazione: (2026)
Advancing Extended Reality with 3D Gaussian Splatting: Innovations and Prospects
di: Qiu, Shi, et al.
Pubblicazione: (2024)
di: Qiu, Shi, et al.
Pubblicazione: (2024)
From Far and Near: Perceptual Evaluation of Crowd Representations Across Levels of Detail
di: Sun, Xiaohan, et al.
Pubblicazione: (2025)
di: Sun, Xiaohan, et al.
Pubblicazione: (2025)
PaintCopilot: Modeling Painting as Autonomous Artistic Continuation
di: Wen, Yunge, et al.
Pubblicazione: (2026)
di: Wen, Yunge, et al.
Pubblicazione: (2026)
Deep Sketch-Based 3D Modeling: A Survey
di: Tono, Alberto, et al.
Pubblicazione: (2026)
di: Tono, Alberto, et al.
Pubblicazione: (2026)
Blind Augmentation: Calibration-free Camera Distortion Model Estimation for Real-time Mixed-reality Consistency
di: Prakash, Siddhant, et al.
Pubblicazione: (2025)
di: Prakash, Siddhant, et al.
Pubblicazione: (2025)
An Evaluation-Centric Paradigm for Scientific Visualization Agents
di: Ai, Kuangshi, et al.
Pubblicazione: (2025)
di: Ai, Kuangshi, et al.
Pubblicazione: (2025)
Vision6D: 3D-to-2D Interactive Visualization and Annotation Tool for 6D Pose Estimation
di: Zhang, Yike, et al.
Pubblicazione: (2025)
di: Zhang, Yike, et al.
Pubblicazione: (2025)
FruitNinja: 3D Object Interior Texture Generation with Gaussian Splatting
di: Wu, Fangyu, et al.
Pubblicazione: (2024)
di: Wu, Fangyu, et al.
Pubblicazione: (2024)
MIBURI: Towards Expressive Interactive Gesture Synthesis
di: Mughal, M. Hamza, et al.
Pubblicazione: (2026)
di: Mughal, M. Hamza, et al.
Pubblicazione: (2026)
3D Reconstruction by Looking: Instantaneous Blind Spot Detector for Indoor SLAM through Mixed Reality
di: Chang, Hanbeom, et al.
Pubblicazione: (2024)
di: Chang, Hanbeom, et al.
Pubblicazione: (2024)
The evolution of volumetric video: A survey of smart transcoding and compression approaches
di: Kakkar, Preetish, et al.
Pubblicazione: (2024)
di: Kakkar, Preetish, et al.
Pubblicazione: (2024)
CleAR: Robust Context-Guided Generative Lighting Estimation for Mobile Augmented Reality
di: Zhao, Yiqin, et al.
Pubblicazione: (2024)
di: Zhao, Yiqin, et al.
Pubblicazione: (2024)
Motion Generation Review: Exploring Deep Learning for Lifelike Animation with Manifold
di: Zhao, Jiayi, et al.
Pubblicazione: (2024)
di: Zhao, Jiayi, et al.
Pubblicazione: (2024)
Improving multidimensional projection quality with user-specific metrics and optimal scaling
di: Ibrahim, Maniru
Pubblicazione: (2024)
di: Ibrahim, Maniru
Pubblicazione: (2024)
Visualization of Age Distributions as Elements of Medical Data-Stories
di: Dowlatabadi, Sophia, et al.
Pubblicazione: (2024)
di: Dowlatabadi, Sophia, et al.
Pubblicazione: (2024)
Virtual Memory for 3D Gaussian Splatting
di: Haberl, Jonathan, et al.
Pubblicazione: (2025)
di: Haberl, Jonathan, et al.
Pubblicazione: (2025)
Glyph-Based Uncertainty Visualization and Analysis of Time-Varying Vector Fields
di: Ouermi, Timbwaoga A. J., et al.
Pubblicazione: (2024)
di: Ouermi, Timbwaoga A. J., et al.
Pubblicazione: (2024)
MRUCT: Mixed Reality Assistance for Acupuncture Guided by Ultrasonic Computed Tomography
di: Wang, Xinkai, et al.
Pubblicazione: (2025)
di: Wang, Xinkai, et al.
Pubblicazione: (2025)
A multidimensional measurement of photorealistic avatar quality of experience
di: Cutler, Ross, et al.
Pubblicazione: (2024)
di: Cutler, Ross, et al.
Pubblicazione: (2024)
Creating Virtual Environments with 3D Gaussian Splatting: A Comparative Study
di: Qiu, Shi, et al.
Pubblicazione: (2025)
di: Qiu, Shi, et al.
Pubblicazione: (2025)
EgoPoseFormer v2: Accurate Egocentric Human Motion Estimation for AR/VR
di: Li, Zhenyu, et al.
Pubblicazione: (2026)
di: Li, Zhenyu, et al.
Pubblicazione: (2026)
NRXR-ID: Two-Factor Authentication (2FA) in VR Using Near-Range Extended Reality and Smartphones
di: Nanzatov, Aiur, et al.
Pubblicazione: (2025)
di: Nanzatov, Aiur, et al.
Pubblicazione: (2025)
Impact of Target and Tool Visualization on Depth Perception and Usability in Optical See-Through AR
di: Yang, Yue, et al.
Pubblicazione: (2025)
di: Yang, Yue, et al.
Pubblicazione: (2025)
SemLayer: Semantic-aware Generative Segmentation and Layer Construction for Abstract Icons
di: Xu, Haiyang, et al.
Pubblicazione: (2026)
di: Xu, Haiyang, et al.
Pubblicazione: (2026)
Beware of Validation by Eye: Visual Validation of Linear Trends in Scatterplots
di: Braun, Daniel, et al.
Pubblicazione: (2024)
di: Braun, Daniel, et al.
Pubblicazione: (2024)
V-Hands: Touchscreen-based Hand Tracking for Remote Whiteboard Interaction
di: Liu, Xinshuang, et al.
Pubblicazione: (2024)
di: Liu, Xinshuang, et al.
Pubblicazione: (2024)
CADReasoner: Iterative Program Editing for CAD Reverse Engineering
di: Kabisov, Soslan, et al.
Pubblicazione: (2026)
di: Kabisov, Soslan, et al.
Pubblicazione: (2026)
A VR Serious Game to Increase Empathy towards Students with Phonological Dyslexia
di: Alcalde-Llergo, José M., et al.
Pubblicazione: (2024)
di: Alcalde-Llergo, José M., et al.
Pubblicazione: (2024)
Advancing Talking Head Generation: A Comprehensive Survey of Multi-Modal Methodologies, Datasets, Evaluation Metrics, and Loss Functions
di: Rakesh, Vineet Kumar, et al.
Pubblicazione: (2025)
di: Rakesh, Vineet Kumar, et al.
Pubblicazione: (2025)
EyeNavGS: A 6-DoF Navigation Dataset and Record-n-Replay Software for Real-World 3DGS Scenes in VR
di: Ding, Zihao, et al.
Pubblicazione: (2025)
di: Ding, Zihao, et al.
Pubblicazione: (2025)
EclipseTouch: Touch Segmentation on Ad Hoc Surfaces using Worn Infrared Shadow Casting
di: Mollyn, Vimal, et al.
Pubblicazione: (2025)
di: Mollyn, Vimal, et al.
Pubblicazione: (2025)
SmartPoser: Arm Pose Estimation with a Smartphone and Smartwatch Using UWB and IMU Data
di: DeVrio, Nathan, et al.
Pubblicazione: (2025)
di: DeVrio, Nathan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
PRAISE: Enhancing Product Descriptions with LLM-Driven Structured Insights
di: Qidwai, Adnan, et al.
Pubblicazione: (2025) -
Unraveling the Truth: Do VLMs really Understand Charts? A Deep Dive into Consistency and Robustness
di: Mukhopadhyay, Srija, et al.
Pubblicazione: (2024) -
ClickAIXR: On-Device Multimodal Vision-Language Interaction with Real-World Objects in Extended Reality
di: Khan, Dawar, et al.
Pubblicazione: (2026) -
MM-Conv: A Multi-modal Conversational Dataset for Virtual Humans
di: Deichler, Anna, et al.
Pubblicazione: (2024) -
Advancements and limitations of LLMs in replicating human color-word associations
di: Fukushima, Makoto, et al.
Pubblicazione: (2024)