Accelerating Physical Property Reasoning for Augmented Visual Cognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Lan, Hongbo, An, Zhenlin, Li, Haoyu, Singh, Vaibhav, Shangguan, Longfei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
por: Ni, Minheng, et al.
Publicado: (2026)
por: Ni, Minheng, et al.
Publicado: (2026)
PoseAugment: Generative Human Pose Data Augmentation with Physical Plausibility for IMU-based Motion Capture
por: Li, Zhuojun, et al.
Publicado: (2024)
por: Li, Zhuojun, et al.
Publicado: (2024)
Visually Grounded Narratives: Reducing Cognitive Burden in Researcher-Participant Interaction
por: Wu, Runtong, et al.
Publicado: (2025)
por: Wu, Runtong, et al.
Publicado: (2025)
Deep Learning in Mild Cognitive Impairment Diagnosis using Eye Movements and Image Content in Visual Memory Tasks
por: Rocha, Tomás Silva Santos, et al.
Publicado: (2025)
por: Rocha, Tomás Silva Santos, et al.
Publicado: (2025)
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
por: Wang, Siting, et al.
Publicado: (2025)
por: Wang, Siting, et al.
Publicado: (2025)
Multimodal LLM Augmented Reasoning for Interpretable Visual Perception Analysis
por: Chaudhari, Shravan, et al.
Publicado: (2025)
por: Chaudhari, Shravan, et al.
Publicado: (2025)
STAR: Smartphone-analogous Typing in Augmented Reality
por: Kim, Taejun, et al.
Publicado: (2025)
por: Kim, Taejun, et al.
Publicado: (2025)
OMG-Bench: A New Challenging Benchmark for Skeleton-based Online Micro Hand Gesture Recognition
por: Chang, Haochen, et al.
Publicado: (2025)
por: Chang, Haochen, et al.
Publicado: (2025)
ConceptFactory: Facilitate 3D Object Knowledge Annotation with Object Conceptualization
por: Sun, Jianhua, et al.
Publicado: (2024)
por: Sun, Jianhua, et al.
Publicado: (2024)
Navigate Biopsy with Ultrasound under Augmented Reality Device: Towards Higher System Performance
por: Li, Haowei, et al.
Publicado: (2024)
por: Li, Haowei, et al.
Publicado: (2024)
Visual Neural Decoding via Improved Visual-EEG Semantic Consistency
por: Chen, Hongzhou, et al.
Publicado: (2024)
por: Chen, Hongzhou, et al.
Publicado: (2024)
End-to-End Motion Capture from Rigid Body Markers with Geodesic Loss
por: Lan, Hai, et al.
Publicado: (2025)
por: Lan, Hai, et al.
Publicado: (2025)
A Monocular SLAM-based Multi-User Positioning System with Image Occlusion in Augmented Reality
por: Lien, Wei-Hsiang, et al.
Publicado: (2024)
por: Lien, Wei-Hsiang, et al.
Publicado: (2024)
OSCAR: Object Status and Contextual Awareness for Recipes to Support Non-Visual Cooking
por: Li, Franklin Mingzhe, et al.
Publicado: (2025)
por: Li, Franklin Mingzhe, et al.
Publicado: (2025)
Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision
por: Li, Chentao, et al.
Publicado: (2026)
por: Li, Chentao, et al.
Publicado: (2026)
Tell Me Without Telling Me: Two-Way Prediction of Visualization Literacy and Visual Attention
por: Chang, Minsuk, et al.
Publicado: (2025)
por: Chang, Minsuk, et al.
Publicado: (2025)
VizDefender: Unmasking Visualization Tampering through Proactive Localization and Intent Inference
por: Song, Sicheng, et al.
Publicado: (2025)
por: Song, Sicheng, et al.
Publicado: (2025)
The Truth, the Whole Truth, and Nothing but the Truth: Automatic Visualization Evaluation from Reconstruction Quality
por: Bujack, Roxana, et al.
Publicado: (2026)
por: Bujack, Roxana, et al.
Publicado: (2026)
Deep Learning-based Lightweight RGB Object Tracking for Augmented Reality Devices
por: Smith, Alice, et al.
Publicado: (2025)
por: Smith, Alice, et al.
Publicado: (2025)
ChildCI Framework: Analysis of Motor and Cognitive Development in Children-Computer Interaction for Age Detection
por: Ruiz-Garcia, Juan Carlos, et al.
Publicado: (2022)
por: Ruiz-Garcia, Juan Carlos, et al.
Publicado: (2022)
VIS-Shepherd: Constructing Critic for LLM-based Data Visualization Generation
por: Pan, Bo, et al.
Publicado: (2025)
por: Pan, Bo, et al.
Publicado: (2025)
RWKV-UI: UI Understanding with Enhanced Perception and Reasoning
por: Yang, Jiaxi, et al.
Publicado: (2025)
por: Yang, Jiaxi, et al.
Publicado: (2025)
SOS: A Shuffle Order Strategy for Data Augmentation in Industrial Human Activity Recognition
por: Ha, Anh Tuan, et al.
Publicado: (2025)
por: Ha, Anh Tuan, et al.
Publicado: (2025)
Toward a Machine Bertin: Why Visualization Needs Design Principles for Machine Cognition
por: Keith-Norambuena, Brian
Publicado: (2026)
por: Keith-Norambuena, Brian
Publicado: (2026)
Adaptive Modality Balanced Online Knowledge Distillation for Brain-Eye-Computer based Dim Object Detection
por: Li, Zixing, et al.
Publicado: (2024)
por: Li, Zixing, et al.
Publicado: (2024)
Spot The Ball: A Benchmark for Visual Social Inference
por: Balamurugan, Neha, et al.
Publicado: (2025)
por: Balamurugan, Neha, et al.
Publicado: (2025)
Panda or not Panda? Understanding Adversarial Attacks with Interactive Visualization
por: You, Yuzhe, et al.
Publicado: (2023)
por: You, Yuzhe, et al.
Publicado: (2023)
Computational Trichromacy Reconstruction: Empowering the Color-Vision Deficient to Recognize Colors Using Augmented Reality
por: Zhu, Yuhao, et al.
Publicado: (2024)
por: Zhu, Yuhao, et al.
Publicado: (2024)
GazeGPT: Augmenting Human Capabilities using Gaze-contingent Contextual AI for Smart Eyewear
por: Konrad, Robert, et al.
Publicado: (2024)
por: Konrad, Robert, et al.
Publicado: (2024)
Dynamics of Affective States During Takeover Requests in Conditionally Automated Driving Among Older Adults with and without Cognitive Impairment
por: Hajian, Gelareh, et al.
Publicado: (2025)
por: Hajian, Gelareh, et al.
Publicado: (2025)
A Comparative Study of Scanpath Models in Graph-Based Visualization
por: Lopez-Cardona, Angela, et al.
Publicado: (2025)
por: Lopez-Cardona, Angela, et al.
Publicado: (2025)
GenColor: Generative Color-Concept Association in Visual Design
por: Hou, Yihan, et al.
Publicado: (2025)
por: Hou, Yihan, et al.
Publicado: (2025)
Enhancing Saliency Prediction in Monitoring Tasks: The Role of Visual Highlights
por: Wu, Zekun, et al.
Publicado: (2024)
por: Wu, Zekun, et al.
Publicado: (2024)
Collection Space Navigator: An Interactive Visualization Interface for Multidimensional Datasets
por: Ohm, Tillmann, et al.
Publicado: (2023)
por: Ohm, Tillmann, et al.
Publicado: (2023)
AIris: An AI-powered Wearable Assistive Device for the Visually Impaired
por: Brilli, Dionysia Danai, et al.
Publicado: (2024)
por: Brilli, Dionysia Danai, et al.
Publicado: (2024)
Self-Supervised Continuous Colormap Recovery from a 2D Scalar Field Visualization without a Legend
por: Liu, Hongxu, et al.
Publicado: (2025)
por: Liu, Hongxu, et al.
Publicado: (2025)
Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models
por: Chen, Tianrun, et al.
Publicado: (2024)
por: Chen, Tianrun, et al.
Publicado: (2024)
Augmented Physics: Creating Interactive and Embedded Physics Simulations from Static Textbook Diagrams
por: Gunturu, Aditya, et al.
Publicado: (2024)
por: Gunturu, Aditya, et al.
Publicado: (2024)
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding
por: Guo, Hao, et al.
Publicado: (2025)
por: Guo, Hao, et al.
Publicado: (2025)
SimVecVis: A Dataset for Enhancing MLLMs in Visualization Understanding
por: Liu, Can, et al.
Publicado: (2025)
por: Liu, Can, et al.
Publicado: (2025)
Ejemplares similares
-
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
por: Ni, Minheng, et al.
Publicado: (2026) -
PoseAugment: Generative Human Pose Data Augmentation with Physical Plausibility for IMU-based Motion Capture
por: Li, Zhuojun, et al.
Publicado: (2024) -
Visually Grounded Narratives: Reducing Cognitive Burden in Researcher-Participant Interaction
por: Wu, Runtong, et al.
Publicado: (2025) -
Deep Learning in Mild Cognitive Impairment Diagnosis using Eye Movements and Image Content in Visual Memory Tasks
por: Rocha, Tomás Silva Santos, et al.
Publicado: (2025) -
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
por: Wang, Siting, et al.
Publicado: (2025)